DevOps / Site Reliability Engineer
Core
Build and operate software systems underpinning uranium enrichment R&D and production infrastructure, ensuring reliability, safety, and developer velocity.
Role type
DevOps / Site Reliability Engineer
Builds
Observability, alerting, CI/CD systems, and internal platforms for nuclear fuel production infrastructure
Domain
Nuclear energy / Heavy industry
Deliverable
infrastructure
Required skills
distributed systems fundamentals, networking (DNS, TLS, HTTP), production system debugging, observability tooling, automation scripting
Preferred skills
modern observability stacks (Prometheus, Grafana, OpenTelemetry), infrastructure-as-code, CI/CD pipelines at scale, safety-critical system support, cross-layer debugging
Technologies
Prometheus, Grafana, OpenTelemetry, Datadog
Responsibilities
Design and maintain observability and alerting systems; ensure production services are instrumented with metrics, logs, and traces; own and maintain developer productivity tools and CI/CD systems; participate in on-call rotation and respond to production incidents; lead incident reviews and drive reliability improvements; automate operational workflows to reduce toil
Seniority
Mid-to-Senior, hands-on IC