Senior Site Reliability Engineer
Core
Build and maintain cloud infrastructure for hosting applications and platforms, while developing internal DevOps tools and best practices.
Role type
Senior Site Reliability Engineer
Builds
Cloud infrastructure, automation tools, and internal DevOps platforms
Domain
Retail technology, Cloud Infrastructure, DevOps
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, Terraform, AWS, CI/CD pipelines, Docker, Bash, Python, authentication and authorization technologies, monitoring and logging
Preferred skills
AI agents, data pipelines (Databricks, Kafka), Java, Javascript, database management (AWS RDS/Postgres), infrastructure cost management
Responsibilities
Manage and scale web applications and data platforms, build tools for DevOps and security best practices, create reusable infrastructure with Terraform, implement monitoring and alerting, develop authentication solutions, investigate application issues, participate in on-call rotations and root cause analysis
Seniority
Senior, hands-on IC