SRE | Site Reliability Engineering
Core
Ensuring product infrastructure health, resilience, and performance while driving migration from EC2 to Kubernetes and automating routine activities.
Role type
Site Reliability Engineer (SRE)
Builds
Cloud infrastructure (AWS, Kubernetes), automated pipelines, and monitoring dashboards for financial products.
Domain
Fintech / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
AWS, Kubernetes, Terraform, Python, Ansible, Rundeck, CI/CD, Shell, PowerShell, Linux, Windows, Grafana, Zabbix, Prometheus, Oracle, PostgreSQL, Redis, MongoDB
Preferred skills
FinOps, incident troubleshooting under pressure, database health management, application security
Technologies
AWS, Kubernetes, Terraform, Rundeck, Python, Ansible, Grafana, Zabbix, Prometheus, Oracle, PostgreSQL, Redis, MongoDB
Responsibilities
Ensure health and resilience of product infrastructure; build metrics, indicators, and dashboards for environment monitoring; identify FinOps opportunities; support new infrastructure creation and scaling; ensure database health; secure applications against vulnerabilities; troubleshoot incidents under pressure.
Seniority
Individual Contributor