Sr Engineer
Core
Architecting, migrating, and optimizing end-to-end monitoring and platform ecosystems for reliability, resilience, and cost-effectiveness.
Role type
Senior IC platform engineer (observability & FinOps)
Builds
Scaled, self-service observability capabilities, standardized frameworks, and automated guardrails for application teams.
Domain
Retail technology, distributed systems, cloud infrastructure
Deliverable
production ML models | infrastructure
Required skills
Observability, Site Reliability Engineering (SRE), FinOps, high-cardinality data management, distributed tracing, SLO/SLI definition, performance optimization, cloud cost showback modeling, automated platform tooling, SDK wrappers, infrastructure abstractions
Preferred skills
Cloud platforms (GCP, AWS, Azure), microservices, DevOps practices, containerization (Docker, Kubernetes), large-scale software systems
Technologies
OpenTelemetry, Golang, Java, Python, C++, Docker, Kubernetes
Responsibilities
Design and maintain opinionated self-service developer paths and observability frameworks; Lead definition and automated enforcement of SLOs and SLIs; Drive strategy for high-cardinality data management and distributed tracing; Implement FinOps observability metrics to enable autonomous cloud footprint management; Collaborate with Finance and Engineering Leadership on cost showback models; Mentor junior and mid-level engineers on performance and FinOps practices; Design scalable observability architectures for long-term business growth; Analyze and resolve complex performance bottlenecks and cost inefficiencies in production.
Seniority
Senior, hands-on IC with mentorship responsibilities