Senior Software Engineer, Observability
Core
Design, build, and own backend systems for metrics, monitoring, and infrastructure maintenance for large-scale server fleets and data centers.
Role type
Senior Software Engineer (Observability/Infrastructure)
Builds
Metrics pipelines, alerting systems, and fleet-wide maintenance/remediation platforms
Domain
Cloud infrastructure, data center engineering, large-scale server fleets
Deliverable
production ML models | product features | infrastructure
Required skills
Python, Go, Linux system debugging, production system design, incident investigation, system reliability
Preferred skills
Ubuntu packaging workflows, CCNA networking experience
Technologies
Python, Go, Linux, Ubuntu
Responsibilities
Design and build services providing visibility into server fleets; Evolve metrics and alerting pipelines; Operate maintenance and remediation systems; Investigate production incidents and drive root-cause fixes; Collaborate with hardware and data center operations teams
Seniority
Senior, hands-on IC