Site Reliability Engineer
Core
Design, build, and monitor scalable, reliable, and fast platforms using SRE principles to guide feature launches and maintain service dependability.
Role type
Site Reliability Engineer (SRE)
Builds
Production infrastructure, CI/CD pipelines, and automated workflows on Azure and Kubernetes
Domain
Fintech / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Public cloud management (Azure), CI/CD tools (Azure DevOps, Octopus, Flux), Linux and Microsoft Systems, Infrastructure as Code (Terraform, Pulumi), Kubernetes and Docker, Scripting (Python, PowerShell, Go), Cloud monitoring (Datadog)
Preferred skills
None explicitly stated as preferred beyond the required list
Technologies
Azure, Datadog, NGINX, Cloudflare, Kubernetes, Serverless, Terraform, CRDs, Crossplane, Azure DevOps, Octopus Deploy, Flux
Responsibilities
Manage and automate Azure, Datadog, NGINX & Cloudflare; Develop and monitor Kubernetes and Serverless resources; Maintain infrastructure code with Terraform & CRDs / Crossplane; Create SLIs and SLOs; Align with Product team on SLAs; Support CI/CD tools; Lead incident troubleshooting
Seniority
Mid-level IC