Site Reliability Engineer
Core
Design, build, and maintain cloud-based infrastructure and services on AWS, ensuring reliability, scalability, and performance.
Role type
Site Reliability Engineer (SRE)
Builds
Cloud infrastructure, CI/CD pipelines, and automation tools
Domain
Cloud infrastructure (AWS)
Deliverable
infrastructure
Required skills
AWS services, Terraform, Python scripting, Bash scripting, incident response, capacity planning, CI/CD pipeline management
Preferred skills
Infrastructure-as-code (IaC) practices, security best practices, zero-downtime release strategies
Technologies
AWS, Terraform, Python, Bash
Responsibilities
Design and maintain cloud infrastructure using Terraform; develop monitoring solutions for performance bottlenecks and outages; participate in incident response and root cause analysis; optimize system performance and cost efficiency; create automation tools for operational tasks; support CI/CD pipelines for application deployments; maintain documentation for infrastructure configurations and processes
Seniority
Mid-level, hands-on IC