Site Reliability Engineer (SRE)
Core
Lead cloud infrastructure team for Apple's Video Computer Vision Organization, ensuring reliability, security, and scalability of ML-based real-time image/video applications.
Role type
Senior Site Reliability Engineer (Infrastructure Lead)
Builds
Cloud-based infrastructure, tooling, and engineering services for Apple's Video Computer Vision applications
Domain
Consumer electronics / Cloud Infrastructure / Machine Learning
Deliverable
infrastructure
Required skills
Kubernetes (EKS/ECS), Terraform, Python, Linux system tuning, AWS services, Docker, Postgres, Redis, Prometheus, hybrid cloud architecture
Preferred skills
Team leadership and mentorship, large-scale production application support, microservices deployment strategies, networking security expertise, kernel tuning
Technologies
Kubernetes, Terraform, Argo, Docker, Python, Postgres, Prometheus, Elasticsearch, Redis, RDS, ELB, AWS
Responsibilities
Develop and maintain infrastructure and tooling for cloud applications, manage system bringup and deployment, drive hybrid infrastructure solutions, provide operational support for security and scalability, mentor team members on best practices
Seniority
Senior, hands-on IC with leadership responsibilities