Global Manager of Site Reliability Engineering (Hybrid - Flexible Options)
Core
Lead the reliability, release engineering, and operational evolution of a global event-driven Integrated Platform connecting enterprise applications with client-facing experiences.
Role type
Senior Manager of Site Reliability Engineering (SRE)
Builds
A high-performing SRE team and a reliable, scalable event-driven platform for Capital Markets and Wealth Management.
Domain
Capital Markets / Wealth Management / Cloud Infrastructure
Required skills
Java, Spring Boot, AWS, Kafka, PostgreSQL, People Leadership, Release Engineering, Infrastructure as Code, CI/CD, Incident Management, Service Level Objectives (via careerplan.io/jobs/JR1085943-global-manager-of-site-reliability-engineering-hybrid-flexible-options-at-broadridge)
Preferred skills
Kubernetes, Terraform, OpenTelemetry, Kafka Connect, PostgreSQL High Availability
Technologies
Java, Spring Boot, AWS, Kafka, PostgreSQL, Kubernetes, Terraform, OpenTelemetry
Responsibilities
Build and lead a high-performing SRE team; Own the SRE strategy and execution roadmap; Remain hands-on in engineering decisions and architecture; Advance AWS infrastructure and platform automation; Improve Kafka reliability and event-processing performance; Strengthen PostgreSQL performance and resilience; Lead release engineering and deployment readiness; Make service health measurable; Reduce operational toil through automation; Enable practical, responsible AI adoption; Lead incident resolution and resilience improvement; Influence cross-functional decisions.
Seniority
Senior, hands-on IC with people management