Site Reliability Engineer II (SRE)
Core
Guide reliability and performance of infrastructure and software systems across engineering teams to enable safe and quick feature deployment.
Role type
Senior IC Site Reliability Engineer
Builds
Reliable and flexible infrastructure solutions for a digital marketplace connecting buyers with manufacturers
Domain
Manufacturing technology / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python, Javascript, Unix Shell, AWS, Terraform, Kubernetes, CI/CD pipelines, Docker
Responsibilities
Develop, configure, and maintain underlying platforms (AWS accounts, networking, kubernetes clusters); Develop, configure, and maintain observability and monitoring tools (Coralogix, Sentry); Develop, configure, and maintain software development tools (github actions runners, ArgoCD); Take ownership of assigned problem statements and drive them to completion; Write clean, efficient, and well-documented code while improving existing systems; Accurately estimate timelines for features and tasks; Participate in an on-call schedule
