Lead Site Reliability Engineer
Core
Lead the full lifecycle of services from architecture through operations, ensuring scalability, reliability, and alignment with business objectives for Mastercard's digital payment products.
Role type
Lead Site Reliability Engineer (SRE)
Builds
Production-ready, fault-tolerant, and scalable payment services and infrastructure
Domain
Fintech / Digital Payments / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Site Reliability Engineering, Infrastructure/DevOps, Linux/UNIX systems, Database administration (Oracle/SQL), Observability/Monitoring, CI/CD pipeline design, Automation/Scripting, Distributed systems design, Cloud platforms (AWS), Program management
Preferred skills
Security environments, System-level design, Cross-functional leadership
Technologies
Splunk, Dynatrace, AWS, Python, Java, Go, C/C++, Perl, Ruby, Oracle, SQL
Responsibilities
Lead service lifecycle from design to optimization; Analyze ITSM performance and influence roadmap prioritization; Define production readiness standards and launch governance; Evolve monitoring frameworks for availability and latency; Champion automation-first principles to reduce toil; Design and govern CI/CD pipelines; Drive incident response practices and postmortems; Mentor junior engineers and raise engineering excellence
Seniority
Senior, hands-on IC with leadership responsibilities