Staff Site Reliability Engineer
Core
Technical lead and architect for the Platform Engineering SRE organization, evolving infrastructure from 'service' to 'product' for the world's leading derivatives marketplace.
Role type
Staff Site Reliability Engineer (Platform Engineering Lead)
Builds
GCP-native stack, self-service internal development platform, and mission-critical Kafka event bus for ultra-low latency financial ecosystems.
Domain
Financial Markets / High-Frequency Trading / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Generative AI & Agentic workflows, Python, Go, GCP (Networking, IAM, GKE), Kafka, Terraform, ArgoCD, Distributed Systems Theory, Executive Communication
Preferred skills
Node.js, modern front-end frameworks, Financial Markets domain expertise
Technologies
GCP, GKE, Kafka, Python, Go, Terraform, ArgoCD, Gemini
Responsibilities
Define 12–18 month technical strategy for Platform SRE teams, act as final technical authority for major infrastructure changes, lead response for complex cross-functional outages, architect high-level abstractions for the internal development platform, standardize SLIs/SLOs/Error Budgets, mentor senior engineers and drive architectural evolution
Seniority
Staff, hands-on IC with strategic leadership