Senior Software Engineer, Site Reliability
Core
Design and build software systems for observability, incident management, problem management, and resiliency to improve reliability and availability across Remitly's engineering organization.
Role type
Senior Software Engineer, Site Reliability Engineering (SRE)
Builds
Scalable SRE tooling and automation for observability, incident response, and system resiliency
Domain
Financial technology / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Software development, Kubernetes, cloud infrastructure (AWS), infrastructure as code, observability tooling, SRE concepts and patterns, technical mentorship
Preferred skills
Leading large ambiguous initiatives, Agile/Kanban methodology
Technologies
Kubernetes, AWS, Prometheus, Grafana, CloudWatch, New Relic, DataDog
Responsibilities
Design and build software projects in SRE domains; Lead large initiatives from inception to delivery; Identify company-wide opportunities to improve availability and implement automation; Provide technical mentorship to other teams on system observation and high availability design
Seniority
Senior, hands-on IC