Senior Site Reliability Engineer
Core
Senior SRE ensuring reliability, performance, and automation for cloud-native financial products and services in Azure environments.
Role type
Senior Site Reliability Engineer (Cloud Native)
Builds
Cloud-native products and services for investment and asset management clients
Domain
FinTech, Cloud Infrastructure, Azure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Azure cloud expertise, Infrastructure as Code (Bicep, ARM, Terraform), monitoring and logging (Azure Monitor, DataDog, Log Analytics), Identity Provider integration (Azure Entra ID, Okta, KeyCloak), observability frameworks (Open Telemetry), synthetic monitoring (Playwright), security tools (Microsoft Defender, Sentinel, KQL), containerization (Kubernetes, Docker), scripting (PowerShell, Bash, Kusto), SQL databases
Preferred skills
AI/ML-based anomaly detection, SimCorp Dimension, Salesforce, ITIL frameworks
Technologies
Microsoft Azure, Bicep, ARM, Terraform, Azure Monitor, Application Insights, DataDog, Log Analytics, Azure Entra ID, Okta, KeyCloak, PingFederate, Open Telemetry, Checkly, Playwright, Microsoft Defender, Sentinel, Kubernetes, Docker, PowerShell, Bash, Kusto, SQL, Cosmos DB, Postgres SQL
Responsibilities
Support operational and enhancement of mission-critical cloud-native environments; Collaborate with product teams to enhance monitoring, observability, and reliability; Manage infrastructure deployment pipelines and troubleshoot onboarding issues; Drive capacity planning for resilient and scalable platforms; Build automation tools to reduce manual toil; Define and manage SLOs and error budgets; Execute disaster recovery and configuration management; Provide on-call support and participate in incident response.
Seniority
Senior, hands-on IC