Senior Observability Engineer
Core
Design and implement observability solutions, monitoring dashboards, and predictive monitoring strategies using AI to ensure system health, performance, and reliability for a financial institution.
Role type
Senior IC Observability Engineer (SRE)
Builds
Production observability platforms, monitoring dashboards, and predictive anomaly detection systems
Domain
Financial Services / Observability & Site Reliability Engineering
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Observability tools (ELK, Dynatrace, Prometheus, Grafana, OpenTelemetry, Jaeger), Python, Java, Go, Query languages, Linux/Unix, Alerting, AI tools for anomaly detection
Preferred skills
Tableau, Financial technology domain, OpenShift, Kubernetes, DevOps/SRE background, Scenario modeling
Technologies
ELK Stack, Dynatrace, Prometheus, Grafana, Open Telemetry, Jaeger, Aternity, Moog, Anthropic Claude, OpenAI, Microsoft Copilot, OpenShift, Kubernetes
Responsibilities
Design and implement observability solutions; Create and maintain monitoring dashboards; Evaluate and recommend observability tools; Analyze logs, metrics, and traces to identify system issues; Develop predictive monitoring strategies using AI; Conduct "what if" analysis using AI capabilities; Mentor junior team members on observability best practices
Seniority
Senior, hands-on IC