Site Reliability Engineer
Core
Provide production support, troubleshooting, and automation for complex Mainframe applications, APIs, and microservices in a banking environment.
Role type
Site Reliability Engineer (Mainframe & API focus)
Builds
Production support for Mainframe (z/OS, CICS, MQ) and modern API/Microservices ecosystems
Domain
Banking / Financial Services / Mainframe & Cloud
Deliverable
production ML models | product features | dashboards & analysis | client delivery | infrastructure | physical/clinical work
Required skills
Mainframe support (z/OS, MVS, CICS, MQ, DB2, COBOL, JCL), Batch scheduling (Zeke, Zena, Control-M), REST/SOAP API support, API modernization, Monitoring (Dynatrace, Splunk, Grafana), Root Cause Analysis, SQL troubleshooting
Preferred skills
API testing tools (Postman, SoapUI), Cloud platforms (AWS, Azure, GCP), ITIL processes, Messaging technologies (Kafka, Event Hub), SRE concepts (SLI/SLO)
Technologies
z/OS, MVS, CICS, MQ, DB2, COBOL, JCL, TSO/ISPF, Zeke, Zena, CA7, Control-M, REST, SOAP, API Gateways, Postman, SoapUI, Swagger/OpenAPI, Dynatrace, Splunk, Grafana, Azure Monitor, AppDynamics, CloudWatch, AWS, Azure, GCP, IBM MQ, Kafka, Event Hub
Responsibilities
Troubleshoot incidents, outages, batch failures, and API performance issues; Support API modernization and SOAP-to-REST migrations; Build monitoring, alerting, and self-healing capabilities; Perform Root Cause Analysis and drive automation; Manage customer-impacting incidents and vendor engagements; Collaborate with Development and Infrastructure teams on deployments and stability.
Seniority
Mid-Senior, hands-on IC