Sr Site Reliability Engineer - Middleware, DevOps & Cloud
Core
Design and operate middleware reliability systems using AI/ML techniques to ensure global payment infrastructure availability and performance.
Role type
Senior Site Reliability Engineer (Middleware & AI/ML)
Builds
Middleware software components, intelligent automation systems, and observability tools for global payment transactions
Domain
Financial Services / Payments / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
DevOps practices, middleware technologies (Tomcat, Apache, Springboot, IBM MQ, etc.), software development (Python, Java, Go), cloud platforms (AWS, GCP, Azure), monitoring and observability (Prometheus, Splunk), Kubernetes, Terraform, Jenkins, Docker, Ansible
Preferred skills
AI & ML engineering, LLM framework integration, open-source contributions
Technologies
Python, Java, Go, Tomcat, Apache, Springboot, MKS, SQS, JBoss, IBM MQ, IBM DataPower, Hazelcast, Flink, Jenkins, Terraform, Ansible, Docker, Kubernetes, AWS, GCP, Azure, Prometheus, Splunk
Responsibilities
Support middleware software and infrastructure components, design AI/ML-based reliability solutions, develop monitoring and observability systems, build intelligent automation tools, troubleshoot production issues and perform root cause analysis, execute middleware releases and deployments, optimize application performance
Seniority
Senior, hands-on IC