Senior Platform Engineer (Observability & Telemetry)
Core
Design, build, and evolve monitoring and observability capabilities to ensure the reliability, performance, and operability of core applications and infrastructure in a financial technology organization.
Role type
Senior IC platform engineer (observability & telemetry)
Builds
Telemetry pipelines, alerting systems, dashboards, and automated remediation workflows for enterprise monitoring
Domain
Financial technology, distributed systems, cloud infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
OpenTelemetry, ElasticStack (APM, Logs, Metrics, Traces), Grafana, OpsRamp, BigPanda, AWS CloudWatch, Azure Monitor, Kubernetes, Docker, Python, Bash, PowerShell, SRE principles, SLI/SLO definition, incident response automation
Preferred skills
CI/CD pipelines (Jenkins, GitHub Actions, GitLab CI), Infrastructure as Code (Terraform, Ansible), microservices architecture, time-series data querying, REST APIs, ServiceNow
Technologies
OpenTelemetry, Elastic, Grafana, OpsRamp, BigPanda, AWS, Azure, Kubernetes, Docker, Jenkins, Terraform, Ansible
Responsibilities
Architect and operate OpenTelemetry-based telemetry pipelines; define and maintain standardized alerting and dashboards; embed reliability considerations (SLOs, error budgets) in the SDLC; build automation frameworks for incident response and self-healing; mentor junior engineers on observability standards; optimize thresholds and escalation processes
Seniority
Senior, hands-on IC