Senior Staff Software Engineer, Tech Foundations
Core
Drive enterprise-wide Site Reliability Engineering (SRE) strategy, architecture, and standards to ensure high availability and performance of Airbnb's infrastructure and products.
Role type
Senior Staff SRE (Strategic IC)
Builds
Enterprise SRE program, reliability architecture, incident management processes, and production readiness standards.
Domain
Travel technology, multi-cloud infrastructure, distributed systems
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
System architecture and design, high-availability system building, multi-cloud environment expertise, incident management, automation, reliable design patterns, mentorship, servant leadership
Preferred skills
Strategic thought partnership, cross-team collaboration, blameless post-mortem facilitation
Technologies
Kubernetes, Terraform, Prometheus, Grafana, AWS, GCP, Azure
Responsibilities
Develop long-term reliability roadmaps and serve as a strategic thought partner; Design and influence company-wide SRE architecture and standards; Create scalable incident management processes; Foster an SRE/Reliability culture with ownership; Develop Production Readiness standards and automate operations; Mentor and lead other Site Reliability Engineers
Seniority
Senior Staff, hands-on IC with strategic influence