Site Reliability Engineer III
Core
Engineer reliability for Google Cloud infrastructure, middleware platforms, and core technology foundations powering Clearing, Risk, and derivatives applications with ultra-low latency and high-concurrency performance.
Role type
Senior Site Reliability Engineer (Platform Engineering)
Builds
Resilient, automated systems and cloud-native middleware platforms for financial derivatives trading
Domain
Financial Markets / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Python, Go, Java, Bash, Linux, Kubernetes, GCP, Terraform, Ansible, TCP/IP, HTTP, DNS, Kafka, Prometheus, Grafana, OpenTelemetry, Splunk
Preferred skills
Generative AI/Agents, Disaster Recovery strategies, Agile frameworks, GCP Professional Cloud Architect, CKA, CKAD, Financial Markets domain expertise
Technologies
Google Cloud Platform, Kubernetes, GKE, Terraform, Ansible, Kafka, RedPanda, Consul, Vault, Splunk, Prometheus, Grafana, OpenTelemetry
Responsibilities
Architect and migrate application platforms to GCP; Design and maintain observability backbone; Engage in live production incidents and lead post-mortems; Eliminate operational toil through automation; Contribute to disaster recovery and resiliency testing; Lead technical discussions and mentor junior colleagues
Seniority
Senior, hands-on IC with mentorship responsibilities