Senior Software Engineer, Site Reliability
Core
Building a new Site Reliability Engineering (SRE) team to ensure the reliability, scalability, and resilience of Asana's distributed systems and infrastructure.
Role type
Senior Software Engineer, Site Reliability Engineering (SRE)
Builds
Reliable distributed systems, internal platforms, frameworks, and incident management processes for Asana's product stack.
Domain
SaaS / Collaboration Platform / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Software engineering, distributed systems design, incident management, system reliability, scalability, infrastructure automation, cloud architecture, on-call rotation management
Preferred skills
Experience with SRE practices, product engineering background, curiosity about AI tools, willingness to learn new technologies
Technologies
AWS, Kubernetes (EKS), Datadog, MySQL (RDS), ElasticSearch (OpenSearch), Redis, DynamoDB, Terraform, TypeScript, Scala, Go, Python
Responsibilities
Influence the future of the SRE practice and shape incident management processes; Lead reliability-focused projects across the stack; Build internal platforms and frameworks to improve service reliability; Participate in a sustainable on-call rotation; Collaborate with global infrastructure teams to solve infrastructure problems
Seniority
Senior, hands-on IC