Distributed Computing DevOps Engineer (IT-CE-LCG-2026-144-LD)
Core
Develop, integrate, and automate operational tools and workflows for a global distributed computing infrastructure supporting high-energy physics research.
Role type
Distributed Computing DevOps Engineer
Builds
Reliable data processing and transfer services for the Worldwide LHC Computing Grid (WLCG) and scientific collaborations.
Domain
High-energy physics / Large-scale distributed computing infrastructure
Deliverable
production ML models | infrastructure
Required skills
Linux-based distributed computing environments, Python programming, web frameworks, relational and non-relational databases, DevOps workflows, Code Management (Jira, GitLab), Continuous Integration/Delivery/Deployment, Containerisation & Orchestration (Kubernetes), monitoring tools (Prometheus, Grafana), large-scale system operations (cloud/grid), data analysis, software lifecycle management, operating systems.
Preferred skills
Experience with large scientific computing projects, working with scientific communities, AI-driven tools and automation technologies.
Technologies
Python, Jira, GitLab, Kubernetes, Prometheus, Grafana, Linux, Cloud/Grid infrastructures
Responsibilities
Develop and automate operational tools and workflows for the HL-LHC programme; monitor and support the operation of a high-scale global distributed computing infrastructure across multiple international sites; evolve and operate monitoring, accounting, and topology systems; support daily operational activities including incident tracking and debugging; contribute to the integration of AI-driven tools and automation technologies.
Seniority
Mid-to-Senior, hands-on IC