Observability Engineer
Core
Design, implement, and optimize enterprise observability solutions across applications, infrastructure, and cloud environments to improve system visibility, operational efficiency, and service reliability.
Role type
Senior IC Observability Engineer
Builds
End-to-end observability solutions, dashboards, alerts, telemetry frameworks, and automation for system health and performance.
Domain
Cloud-native infrastructure and distributed systems
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Observability platforms (Dynatrace, Splunk, Grafana, OpenTelemetry), AWS, GCP, Python, Terraform, SLIs/SLOs, incident response frameworks, distributed tracing (MELT), CI/CD integration, ITSM tools (ServiceNow)
Preferred skills
AIOps platforms, Kubernetes, observability tool integration with CI/CD and ServiceNow, AWS/GCP/Observability/SRE certifications
Technologies
Dynatrace, Splunk, Grafana, OpenTelemetry, AWS, GCP, Python, Terraform, ServiceNow, Kubernetes
Responsibilities
Design and implement end-to-end observability solutions; Develop dashboards, alerts, and telemetry frameworks; Build automation solutions for operational tasks; Enable runbook automation and self-healing capabilities; Define and implement SLIs, SLOs, and alerting strategies; Mentor teams and drive adoption of observability best practices.
Seniority
Senior, hands-on IC