Expert Observability Engineer
Core
Architect and govern unified observability frameworks, lead SRE incident management, and drive platform engineering automation for enterprise clients.
Role type
Senior IC Observability Engineer / Lead Architect
Builds
Production-ready observability patterns, telemetry pipelines, and automated monitoring workflows for multi-cloud environments.
Domain
IT Services / Observability / Cloud-Native Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Enterprise Architecture, SRE practices, Observability-as-Code, Incident Management, Cloud-Native Design, Automation, Vendor Transition Leadership
Preferred skills
Kubernetes Administration, Cloud Architecture, APM Vendor Certifications
Technologies
IBM Instana, Grafana, OpenTelemetry, Telegraf, InfluxDB, Prometheus, SolarWinds, Netcool, Elastic, Splunk, Kubernetes, Docker, OpenShift, AWS, Azure, GCP, Linux, RHEL, Windows Server, VMware, Citrix, Ansible, Terraform, Python, Bash, ServiceNow, ITIL 4, GitHub Actions, GitLab, Jenkins
Responsibilities
Architect and govern unified observability frameworks covering metrics, logs, traces, and events; Define and govern Service Level Indicators (SLIs), Objectives (SLOs), and error budgets; Drive Observability-as-Code and infrastructure automation; Design deep observability for Docker, Kubernetes, microservices, and multi-cloud environments; Lead complex Knowledge Transfer (KT) programs and mentor cross-functional engineering teams.
Seniority
Senior, hands-on IC with leadership responsibilities