Senior Site Reliability Engineer
Core
Production readiness owner ensuring systems are designed, built, and operated with reliability, scalability, and performance at their core.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Platform services across the software lifecycle from design to production
Domain
Financial technology / Payments infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
UNIX/Linux systems administration, scripting and automation, CI/CD pipeline management, capacity planning, observability and monitoring, incident response, system design consulting
Preferred skills
Knowledge of Artificial Intelligence use cases and implementation, experience with C/C++/Java/Python/Go/Perl/Ruby
Technologies
Splunk, Dynatrace, Oracle, SQL
Responsibilities
Engage in and improve the whole lifecycle of services from inception to refinement; Analyze ITSM activities and provide feedback on operational gaps; Support services pre-launch through system design consulting and capacity planning; Maintain live services by measuring availability, latency, and system health; Scale systems sustainably through automation; Practice sustainable incident response and blameless postmortems; Mentor junior resources
Seniority
Senior, hands-on IC
