SRE, London
Core
Design, engineer, and run large-scale distributed database infrastructure (FoundationDB) for Apple Services, ensuring global availability and performance for hundreds of millions of users.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Production-grade distributed database systems (FoundationDB) powering Apple Services like Mail, Contacts, and Keychain
Domain
Cloud infrastructure, distributed systems, database operations
Required skills
Distributed systems management, Linux OS, Go, Java, Python, Kubernetes, Configuration management (Puppet, Chef, Ansible), Capacity planning, Disaster recovery, Microservices architecture
Preferred skills
Kubernetes operator development, Scale testing, Automation design, Troubleshooting complex infrastructure issues
Technologies
FoundationDB, Kubernetes, Linux, AWS, Java, Go, Python, Puppet, Chef, Ansible, Spinnaker
Responsibilities
Provision, manage, and monitor FoundationDB across multiple regions and control planes (bare-metal, AWS, Kubernetes); Develop automation tools in Java and Go; Collaborate with development teams to deliver robust and scalable database solutions; Perform scale testing, disaster recovery planning, and capacity planning.
Seniority
Senior, hands-on IC