Senior Site Reliability Engineer
Core
Oversee a group of employees to ensure operational goals are met and services run smoothly while managing the full lifecycle of services.
Role type
Senior Site Reliability Engineer (Team Lead)
Builds
Production services, automated tooling, and scalable infrastructure for online sports betting and iGaming platforms.
Domain
iGaming / Online Sports Betting / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Team management and mentoring, incident management, SRE principles, containerized microservices, CI/CD pipelines, distributed systems monitoring, capacity planning, performance testing, automation, networking, observability, programming (Java/Python/Scala/Kotlin), technical decision making, agile/scrum leadership.
Preferred skills
Managing complex telemetry solutions, Infrastructure as Code (Chef/Puppet/Terraform), edge configuration (CDN/certificates), SLO culture building, industry thought leadership.
Technologies
Jenkins, Buildkite, GitHub Actions, AWS, GCP, Pagerduty, Opsgenie, Java, Python, Scala, Kotlin, Chef, Puppet, Terraform, CDN.
Responsibilities
Oversee a group of employees and provide direction for operational goals, engage in the full lifecycle of services (design to refinement), investigate and resolve production problems, support pre-launch activities (design consulting, capacity planning), monitor live services for availability and latency, perform performance and capacity testing, optimize reliability monitoring and alerting, scale systems via automation, conduct performance auditing and define SLIs, practice blameless postmortems, lead projects in agile environments, act as on-call for services.
Seniority
Senior, hands-on IC with team leadership responsibilities