Senior Site Reliability Engineer
Core
Design, build, and maintain highly scalable and reliable cloud-native infrastructure for a globally distributed, highly available system.
Role type
Senior Site Reliability Engineer (IC)
Builds
Production infrastructure, automation tools, and observability solutions for Symphony's communication platform.
Domain
Cloud infrastructure, DevOps, and SRE
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Infrastructure-as-Code (IaC), Kubernetes, Linux administration, Cloud platforms (GCP, AWS), Observability, Python, Go, Networking
Preferred skills
Terragrunt, Ansible, Helm, ArgoCD, Splunk, Grafana
Technologies
Terraform, Kubernetes, Helm, ArgoCD, GCP, AWS, Splunk, Grafana, Python, Go
Responsibilities
Design and maintain scalable infrastructure using IaC; Manage and optimize Kubernetes clusters; Administer and troubleshoot Linux-based systems; Architect and manage cloud-native solutions on GCP and AWS; Implement robust observability practices; Develop and maintain automation tools; Provide production support and participate in 24/7 on-call rotation; Troubleshoot complex networking issues.
Seniority
Senior, hands-on IC