Software Engineer Spec III (SRE)
Core
Senior Site Reliability Engineer responsible for sustaining and evolving cloud and on-premises infrastructure, ensuring system availability, and supporting development teams with CI/CD pipelines.
Role type
Senior SRE (Infrastructure & Cloud)
Builds
Cloud infrastructure (AWS), on-premises environments (Linux/Windows), CI/CD pipelines (Azure DevOps), and observability stacks.
Domain
Financial Services / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Infrastructure as Code (Terraform, CloudFormation), Linux/Windows administration, CI/CD (Azure DevOps), Containers & Orchestration (Docker, Kubernetes), Observability (Zabbix, Grafana, Prometheus, ELK), Networking (VPC, DNS, Subnets), IAM Policies & Roles.
Preferred skills
AWS Serverless services (Lambda, API Gateway, CloudFront, S3), APM (Dynatrace), Agile/DevOps methodologies, Scripting (Python, Bash, PowerShell), AI tools.
Responsibilities
Manage and evolve cloud and on-premises infrastructure; Support development teams in creating and optimizing CI/CD pipelines; Monitor systems proactively and manage alerts; Troubleshoot critical incidents and perform root cause analysis; Implement observability metrics and dashboards; Facilitate production environment requests across teams.
Seniority
Senior, hands-on IC