Infrastructure engineer (UK)
Core
Build resilient, high-availability infrastructure and automation for an enterprise generative AI platform serving hundreds of companies.
Role type
Senior Infrastructure Engineer (SRE/DevOps/Platform)
Builds
End-to-end production infrastructure, release pipelines, and agentic workflows for enterprise AI agents.
Domain
Enterprise Generative AI / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, Terraform, Python, Go, AWS, GCP, Azure, Prometheus, Grafana, ELK, Incident Response, Root Cause Analysis, SLO management, Infrastructure as Code, Containerization, AI-assisted workflows
Preferred skills
Software engineering background, 0-to-1 infrastructure builds, Cross-functional collaboration
Technologies
Kubernetes, Helm, Terraform, Pulumi, AWS, GCP, Azure, Python, Go, Claude Code, Droid, Codex, Prometheus, Grafana, ELK
Responsibilities
Design scalable fault-tolerant infrastructure across cloud providers; Automate operational tasks and infrastructure management; Lead incident response and post-mortems; Balance tactical fixes with long-term platform direction; Operate AI agents in daily workflows to investigate incidents and draft changes.
Seniority
Senior, hands-on IC