Staff Engineer
Core
Design, build, and operate the internal and external Observability stack for the MongoDB platform, handling tens of billions of metrics time series and petabytes of logs/traces to safeguard critical customer workloads.
Role type
Staff Engineer (Distributed Systems & Observability Infrastructure)
Builds
Mission-critical observability systems (metrics, logs, traces, alerts) for MongoDB Atlas and internal engineering teams.
Domain
Database infrastructure, distributed systems, observability.
Deliverable
production ML models | product features | infrastructure
Required skills
Designing and tuning distributed/highly concurrent C/C++/Java/Rust systems, multi-threaded programming, performance profiling, database internals, observability ecosystem expertise, mentoring engineers, technical leadership, system architecture.
Preferred skills
Indexing or database performance tuning experience, familiarity with VictoriaMetrics, Splunk, Flink, WarpStream/Kafka, Java, Golang, Fluentbit.
Technologies
VictoriaMetrics, Splunk, Flink, WarpStream, Kafka, Java, Golang, Fluentbit
Responsibilities
Architect and implement components for the observability platform, design improvements for root cause diagnosis, handle production customer escalations, write and review production database code, diagnose test failures and fix bugs, investigate performance regressions, interview candidates, lead large projects, advise Product Management on technical direction, collaborate on roadmaps, champion unit/integration tests, participate in on-call rotation.
Seniority
Staff, hands-on IC with technical leadership