Site Reliability Engineer
Core
Ensure Service Level Agreements, monitor production environments for scalability and security, and build software to automate manual operational tasks.
Role type
Site Reliability Engineer
Builds
Business critical services and global product platform
Domain
Healthcare technology / Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Azure, Kubernetes, Python, Go, MySQL, Terraform, Unit testing, Integration testing, Static analysis, Resilience testing, Root cause analysis, On-call support
Preferred skills
Security-first design, Best practices engineering, Complex problem-solving
Technologies
Azure, Kubernetes, Python, Go, MySQL, Terraform
Responsibilities
Respond to incidents and perform root cause analysis; Build automation tools to reduce manual tasks; Design and scale global product platform; Provide out-of-hours on-call support; Propose infrastructure improvements.
Seniority
Individual Contributor