Senior SRE, Software Engineering
Core
Ensure reliability, scalability, and performance of production and internal systems for a consumer finance platform scaling from thousands to millions of borrowers.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Resilient infrastructure, observability solutions, and sustainable on-call processes for high-availability systems.
Domain
Consumer Finance / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Incident response leadership, root cause analysis, highly available deployment architecture, monitoring & observability design, AWS cloud services, infrastructure as code, CI/CD pipeline design
Preferred skills
Blameless postmortem culture, async workflow infrastructure, database optimization, data pipeline reliability
Technologies
Terraform, CloudFormation, AWS, Datadog, Prometheus, ELK
Responsibilities
Lead incident response and establish sustainable on-call practices; Develop self-service observability solutions; Create and maintain infrastructure as code; Partner with feature teams to architect resilient infrastructure; Work with DevX to design CI/CD pipelines; Advocate for reliability best practices in feature design
Seniority
Senior, hands-on IC
