Site Reliability Engineer
Core
Providing technical and production support for a large-scale global financial software product, ensuring system reliability, scalability, and performance with 24x7 coverage.
Role type
Site Reliability Engineer (SRE)
Builds
Financial software applications and services on AWS and Azure platforms
Domain
Financial services / Cloud Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Cloud computing (AWS/Azure), scripting (Python/PowerShell/Bash), system administration (Windows/Linux), CI/CD, observability, incident response, automation
Preferred skills
Azure/AWS certifications, Power Platform, LLM integration, Financial services domain experience
Technologies
Azure (AKS, ACR, Functions, Web App, Containers, LLM), AWS (RDS, CloudFront, DynamoDB, Lambda, CloudWatch, Kinesis), Git, GitLab, ServiceNow, Confluence, VS Code, Copilot
Responsibilities
Maintain software applications for scalability and reliability; monitor service availability and latency; implement DevOps practices and manage release processes; provide 24x7 client-facing technical support; lead incident response and problem resolution; mentor team members on automation and service performance
Seniority
Senior Associate