Infrastructure Engineer - Observability (APAC)
Core
Design, operate, and scale high-throughput observability systems (logging, tracing, metrics) to ensure high availability for internal and external stakeholders.
Role type
Senior Infrastructure Engineer (Observability)
Builds
Production observability platforms (logging, tracing, metrics) used by internal and external customers
Domain
Cloud Infrastructure / Observability
Deliverable
production ML models | product features | infrastructure
Required skills
Kubernetes administration, Infrastructure as Code (Terraform/Pulumi/OpenTofu), PostgreSQL administration, telemetry tooling (OpenTelemetry, VictoriaMetrics, Grafana, Prometheus), high-throughput data pipeline orchestration
Preferred skills
AWS services experience, high availability and systems reliability in high volume environments
Technologies
Kubernetes, VictoriaMetrics, OpenTelemetry, Vector, Grafana, Prometheus, PostgreSQL, Terraform, Pulumi, OpenTofu, AWS
Responsibilities
Collaborate with infrastructure and product teams to enforce org-wide telemetry practices; Own and operate Kubernetes infrastructure for the observability team; Develop reliability software to ensure observability systems uptime; Orchestrate and scale systems like VictoriaMetrics, OpenTelemetry Collector, and Vector
Seniority
Senior, hands-on IC