Site Reliability Engineer
Core
Perfect enterprise infrastructure DevOps practices and run hundreds of private cloud, Kubernetes clusters, and applications across physical and public cloud estates.
Role type
Site Reliability Engineer (Infrastructure)
Builds
Private cloud, Kubernetes clusters, storage solutions, and open source applications
Domain
Open source infrastructure, cloud computing, DevOps
Deliverable
production ML models | product features | infrastructure
Required skills
Linux, Python, networking, cloud architecture, Kubernetes operations, bare-metal networking, kernel knowledge
Preferred skills
OpenStack deployment, public cloud deployment, private cloud management
Technologies
OpenStack, Kubernetes, Kubeflow, Kafka, OpenSearch
Responsibilities
Deploy and run OpenStack, Kubernetes, storage solutions, and open source applications; identify and address incidents; monitor and observe applications; anticipate potential issues; enable product refinement
Seniority
Mid-level, hands-on IC