Senior Systems Reliability Engineer
Core
Build and operate high-quality production systems for film production and direct-to-consumer mobile applications, ensuring high availability and rapid development.
Role type
Senior Systems Reliability Engineer (IC)
Builds
Cloud-native film making systems and infrastructure for Disney studios
Domain
Media production technology / Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
System management languages (Terraform, Ansible), Public cloud platforms (AWS, Google, Azure), Scripting (Python, GO, Ruby, Swift), Containerization (Docker, Kubernetes), CI/CD tools (Jenkins, GitLab CI/CD), Observability tools (DataDog, New Relic, Grafana), Operating systems (Amazon Linux, Windows), Authentication tools (Active Directory, LDAP, Ping Identity), Data center and network architecture, Systems security (key management, encryption, vulnerability management)
Preferred skills
Agentic coding tools (Claude Code, Cursor, Github Copilot), Media production environment experience, Virtual hosting technologies (VMWare, KVM)
Technologies
Terraform, Ansible, AWS, Google Cloud, Azure, Python, GO, Ruby, Swift, Docker, Kubernetes, Jenkins, GitLab CI/CD, DataDog, New Relic, Grafana, Active Directory, LDAP, Ping Identity, Amazon Linux, Windows, VMWare, KVM
Responsibilities
Build and operate high-quality production systems; Design systems to enable rapid development, high availability, and clear observability; Maintain and improve the reliability of services and infrastructure; Troubleshoot and resolve performance and reliability issues across the stack; Create self-healing infrastructure-as-Code and automate everything; Ensure security best practices in technical solutions
Seniority
Senior, hands-on IC