Sr Engineer, SRE TechOps CICD (Remote)
Core
Senior SRE owning the availability, health, and automation of CrowdStrike's internal CICD environment and developer platform.
Role type
Senior Site Reliability Engineer (Technical Operations)
Builds
Internal developer platform, CICD pipelines, and operational automation for thousands of engineers.
Domain
Cybersecurity / Cloud Infrastructure / DevOps
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, CI/CD tools (Bazel, Jenkins, GitLab CI, GitHub Actions), IaC (Terraform, Ansible, Chef, Puppet, Salt), Observability (Datadog, Grafana, Prometheus, Splunk), Databases (Postgres, MySQL, MongoDB, Cassandra, Kafka, Redis), Scripting (Python, Go, Bash), Big Data systems, Security principles
Preferred skills
AI-assisted workflows, Networking (Load balancers, DNS, Firewalls), Hybrid cloud/on-prem environments, Data science/ML principles
Technologies
Kubernetes, Bazel, Jenkins, GitLab CI, GitHub Actions, Terraform, Ansible, Chef, Puppet, Salt, Datadog, Grafana, Humio, New Relic, Prometheus, Splunk, Cassandra, Postgres, MySQL, MongoDB, OpenSearch, Kafka, Redis, Valkey, Python, Go, Bash, PowerShell, AWS, Azure, GCP, Oracle, Apache Airflow, Apache Spark
Responsibilities
Own service availability and health within the CICD environment; Build software and systems for platform infrastructure and deployment automation; Carry on-call responsibility and drive incident response/postmortems; Gather and analyze metrics for performance tuning and root cause analysis; Lead system design discussions, production readiness reviews, and capacity planning; Evaluate and integrate agentic/AI-assisted workflows into team processes; Mentor mid-level and junior engineers; Investigate emerging technologies and provide roadmap recommendations; Build automated reporting on service health and compliance; Partner with peers to drive cross-team reliability improvements.
Seniority
Senior, hands-on IC with mentorship responsibilities