Observability Engineer
Core
Design, implement, and operate an end-to-end observability stack (logging, metrics, monitoring, alerting, tracing) for cloud environments to ensure system visibility and performance.
Role type
Senior IC Observability Engineer
Builds
Robust observability infrastructure and automated alerting systems for cloud-native workloads
Domain
Cloud Infrastructure & Observability
Deliverable
production ML models | infrastructure
Required skills
Infrastructure as Code (Terraform), Python, Kubernetes, Cloud Platforms (AWS/Azure/GCP/OCI), Observability Tools (Prometheus/Grafana/Loki/PagerDuty), CI/CD pipelines, End-to-end monitoring design
Preferred skills
Application analysis, Cross-team integration with DevOps/SRE
Technologies
Terraform, Python, Kubernetes, AWS, Azure, OCI, GCP, Prometheus, Grafana, Loki, PagerDuty, Jenkins, GitLab, GitHub
Responsibilities
Design and implement observability stack components; Automate provisioning using IaC; Define organization-wide monitoring standards; Develop automated alerting systems; Collaborate with DevOps and SRE teams to align strategies
Seniority
Senior, hands-on IC