Senior Observability Infrastructure Engineer
Core
Designing and operating large-scale observability infrastructure (logs, metrics, traces) for hundreds of product teams, managing petabytes of data across hybrid on-premise and Kubernetes environments.
Role type
Senior IC Observability Infrastructure Engineer
Builds
Next-generation logging and metrics systems, self-healing automation platforms, and global infrastructure supporting new regions.
Domain
Financial Technology / Platform Engineering / Distributed Systems
Deliverable
production ML models | product features | infrastructure
Required skills
Linux kernel-level debugging, Kubernetes operations (on-prem/cloud), Go or Python software engineering, Elasticsearch/OpenSearch/VictoriaLogs/Clickhouse, Prometheus/VictoriaMetrics, Grafana Tempo, OpenTelemetry, CI/CD pipeline automation, performance tuning of high-volume data streams.
Preferred skills
Multi-tenant isolation and quota governance, experience in regulated environments.
Technologies
Go, Python, Kubernetes, Elasticsearch, OpenSearch, VictoriaLogs, Clickhouse, Prometheus, VictoriaMetrics, Grafana, Tempo, OpenTelemetry, kubectl.
Responsibilities
Design and implement future architecture for logging and metrics systems; manage lifecycle of 1,500+ servers across bare-metal and Kubernetes; write code to automate operational tasks and build self-healing systems; tune performance bottlenecks in distributed tracing and logging pipelines; participate in on-call rotations while engineering solutions to prevent alerts.
Seniority
Senior, hands-on IC