Site Reliability Engineer
Core
Perfect enterprise infrastructure DevOps practices and raise the bar on automation using a model-driven approach across on-premise and public clouds.
Role type
Site Reliability Engineer (SRE)
Builds
Private cloud, Kubernetes clusters, and applications for global enterprise customers
Domain
Open source infrastructure, cloud computing, DevOps
Deliverable
production ML models | infrastructure
Required skills
Python software development, Linux operations, Kubernetes deployment/operations, networking, cloud architecture
Preferred skills
OpenStack deployment/operations, public cloud management, private cloud management
Technologies
OpenStack, Kubernetes, Kubeflow, Kafka, OpenSearch, databases
Responsibilities
Deploy and run OpenStack, Kubernetes, storage solutions, and open source applications; identify and address incidents; monitor and observe applications; anticipate potential issues
Seniority
Mid-level to Senior, hands-on IC
