Observability Engineer
Core
Building, scaling, and embedding enterprise-grade observability across the technology landscape to enable reliable, secure, and high-performing services.
Role type
Lead Observability Engineer
Builds
Enterprise observability platform covering logs, metrics, traces, APM, RUM, monitoring, and security use cases
Domain
Financial technology / Cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Datadog (logs, metrics, APM, dashboards, monitors), Python (automation, instrumentation), AWS monitoring services, Azure monitoring services, DevOps tooling, CI/CD pipelines, Infrastructure-as-code (Terraform), observability principles (telemetry types, SLOs, distributed tracing, reliability metrics)
Preferred skills
Kubernetes, container-based observability patterns, OpenTelemetry instrumentation, SRE background, production operations, performance engineering
Technologies
Datadog, Python, AWS, Azure, Terraform, Kubernetes, OpenTelemetry
Responsibilities
Define and implement reusable observability standards for dashboards, alerting, and logging; Embed observability into cloud infrastructure and CI/CD pipelines; Partner with engineering, SRE, security, and operations teams to improve incident detection and service reliability; Enable self-service observability through documentation and dashboards
Seniority
Senior, hands-on IC