Member of Technical Staff - Observability
Core
Building and operating the core infrastructure for monitoring, debugging, and optimizing system performance and reliability at massive scale.
Role type
Senior IC infrastructure engineer (observability)
Builds
High-performance telemetry pipelines, metrics, logs, tracing, and alerting systems for engineering teams.
Domain
Cloud infrastructure, distributed systems, observability
Deliverable
production ML models | infrastructure
Required skills
Go, Rust, Scala, distributed systems, telemetry architecture, large-scale infrastructure, observability stacks, Kafka, Redis, time series databases, Kubernetes
Preferred skills
Prometheus, Grafana, OpenTelemetry, VictoriaMetrics, ClickHouse, API design, UI development
Responsibilities
Design and implement scalable observability infrastructure; Build high-performance telemetry pipelines; Develop APIs, query engines, and UIs; Define and enforce best practices for instrumentation and reliability; Partner with infrastructure and product teams; Own the reliability, scalability, and performance of the observability stack.
Seniority
Senior, hands-on IC