Sr. Site Reliability Engineer - Digital Assets
Core
Apply software and systems engineering to improve reliability, resilience, scalability, and operational health of production services.
Role type
Senior Site Reliability Engineer
Builds
Production services, CI/CD pipelines, observability tooling, and automation for financial payment systems
Domain
Financial services / Payments / Cloud Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Software engineering, distributed systems, automation, observability, incident response, capacity management, Linux/Unix, public cloud architecture
Preferred skills
AWS, containers/orchestration, Infrastructure as Code, SLIs/SLOs/error budgets, disaster recovery, reusable tooling development
Technologies
AWS, Azure, GCP, OCI, Linux/Unix, CI/CD, containers, IaC
Responsibilities
Define and implement SLIs, SLOs, and error budgets; lead incident response and post-incident learning; reduce operational toil through automation; partner with engineering teams to embed reliability into the development lifecycle; mentor engineers and share reusable practices
Seniority
Senior, hands-on IC with mentorship responsibilities
