Sr. Staff Engineer Software, Infrastructure Reliability (Chronosphere)
Core
Architect and build high-scale developer tooling and backend services to enable developer velocity and reliability for the entire engineering organization.
Role type
Sr. Staff Engineer, Infrastructure Reliability
Builds
End-to-end development lifecycle platform including local tools, CI/CD, dynamic testing environments, and deployments for observability.
Domain
Cloud-native infrastructure and developer productivity
Deliverable
production ML models | product features | infrastructure
Required skills
Go/Java/Python/Rust, Terraform, Kubernetes, Linux internals, distributed systems, networking, security, debugging, system design
Preferred skills
AI coding assistants, cloud-native concepts, ephemeral infrastructure
Technologies
Terraform, Kubernetes, AWS, GCP
Responsibilities
Design and maintain high-scale developer tooling and backend services; define and manage environments using declarative IaC; identify and eliminate systemic bottlenecks in the SDLC; ensure infrastructure resilience under massive traffic loads; define platform standards and reference architectures; mentor teams on infrastructure best practices.
Seniority
Sr. Staff, strategic leadership & hands-on IC