Principal Engineer - Observability Telemetry Client Infrastructure
Core
Architect and build a self-service, multi-tenant OpenTelemetry-based telemetry fabric for high-throughput log, metric, and trace ingestion and correlation in massive-scale OLAP engines.
Role type
Principal Engineer (Systems Architecture & Infrastructure)
Builds
High-performance telemetry SDKs, automated client on-ramps, and unified trace stitching logic for hybrid-cloud environments.
Domain
Cloud-native observability, distributed systems, high-cardinality data processing
Deliverable
production ML models | product features | infrastructure
Required skills
Large-scale distributed systems architecture, OpenTelemetry Collector design, high-volume OLAP engine optimization, multi-language SDK development, semantic data modeling, cross-organizational technical leadership
Preferred skills
Kubernetes-native observability, CNCF project contributions, AIOps/advanced analytics, industry conference speaking
Technologies
OpenTelemetry, ClickHouse, StarRocks, Java, Go, Python, Kubernetes
Responsibilities
Architect automated on-ramps for multi-cloud/hybrid observability clients; enforce semantic conventions for telemetry correlation; develop high-performance interfaces and SDKs; stitch disparate signals into unified traces; conduct deep-dive code reviews and resolve systemic bottlenecks; mentor senior/staff engineers.
Seniority
Principal, hands-on IC with strategic influence