Sr Site Reliability Engineer, Customer Systems , IS&T
Core
Design, build, and deliver highly scalable, reliable, secure cloud infrastructure powering Apple's customer applications and services.
Role type
Senior Site Reliability Engineer (Customer Systems)
Builds
Highly available, scalable, reliable, secure Infrastructure
Domain
Cloud Infrastructure, Customer Systems
Required skills
Kubernetes (Helm), Shell Scripting, Python, Ansible, Splunk, Grafana, Prometheus, Alertmanager, DNS, TCP, HTTP/HTTPS, CI/CD pipelines
Preferred skills
Cassandra, MongoDB, Couchbase, AWS S3, Java application support, ArgoCD, GitOps, MTTR/SLO metrics, GenAI tools for workflow automation
Technologies
Kubernetes, Helm, Splunk, Grafana, Prometheus, Alertmanager, Cassandra, MongoDB, Couchbase, AWS S3, ArgoCD, GenAI
Responsibilities
Innovate, architect, build, and document highly available, scalable, reliable, secure Infrastructure; Troubleshoot application specific, network, system & performance issues; Build and maintain CI/CD infrastructure; Envision and build automation tools; Collaborate with other site reliability engineers, software engineers, quality engineers to gather, define, and analyze non-functional/technical requirements
Seniority
Senior, hands-on IC
