Lead Site Reliability Engineer
Core
Lead strategic and technical ownership of cloud infrastructure, service reliability, and platform operations for a FinTech product area, ensuring excellence for newly onboarded and long-standing clients.
Role type
Lead Site Reliability Engineer (Cloud Infrastructure)
Builds
Azure-hosted environments, automated scalable solutions, observability frameworks, and disaster recovery plans.
Domain
Financial Technology (FinTech) / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Azure cloud leadership, Site Reliability Engineering (SRE) practices, Infrastructure as Code (Terraform, Bicep, ARM, Ansible), incident response and management, Kubernetes, Docker, CI/CD, SQL, ITIL practices, mentoring and technical leadership.
Preferred skills
Experience with SimCorp Dimension or financial services platforms.
Technologies
Microsoft Azure, Terraform, Bicep, ARM, Ansible, Kubernetes, Docker, SQL
Responsibilities
Own reliability, scalability, and performance of Azure-hosted environments; lead operational support and onboarding for client platforms; provide technical leadership in SRE practices and incident response; guide transition of manual processes to automation; implement observability frameworks using SLOs/SLIs; oversee disaster recovery and postmortems; mentor engineers; collaborate with product owners on platform design; provide on-call support.
Seniority
Senior, hands-on IC with leadership responsibilities