Senior Site Reliability Engineer
Core
Designing, implementing, and troubleshooting complex network infrastructures and cloud-native platforms to ensure high availability and reliability for electronic trading systems.
Role type
Senior Site Reliability Engineer (IC)
Builds
Highly available, secure, and reliable cloud-native platforms for electronic trading
Domain
Financial Services / Cloud Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Kubernetes, AWS (Lambda, EKS, SNS), Python, Linux/Unix, GitSecOps, ArgoCD, Pulumi, Kustomize, Network Architecture, Threat Modeling, Observability, SLO definition
Preferred skills
LGTM, Maven, Web DevOps
Technologies
AWS Lambda, ArgoCD, Kubernetes, Linux, Python, Maven
Responsibilities
Automate delivery through GitSecOps practices to reduce toil and increase reliability; Triage incidents, assess risk, and own production infrastructure issues until resolved; Build and maintain monitoring and observability tools to support SLOs; Collaborate with software teams to shape architecture for reliability and performance; Participate in Agile planning and story ideation for platform improvements; Perform on-call duties to maintain high system availability.
Seniority
Senior, hands-on IC