Senior Site Reliability Engineer
Core
Proactively and reactively improve the reliability of Block's platform and critical infrastructure using AI-driven tooling and automation.
Role type
Senior Site Reliability Engineer (IC)
Builds
Distributed platforms enabling safe, scalable product development for Block's financial services and crypto products.
Domain
Financial technology / Cryptocurrency / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Incident command and triage, AI-driven observability and automation, CI/CD pipeline management, progressive delivery strategies, root cause analysis, vendor dependency management, evidence-based maturity assessments
Preferred skills
Experience with backend/platform focus projects, autonomy in high-availability systems
Technologies
Kotlin, Modern Java (11+), HTTP, JSON, gRPC, Protocol Buffers, MySQL, Vitess, DynamoDB, Event driven architectures, DataDog, LaunchDarkly, Terraform, Kubernetes, Istio, Envoy, Amazon Web Services
Responsibilities
Lead incident command and coordinate mitigation for Tier 0 services, build and extend platforms to improve system reliability, standardize reliability tools across multiple platforms, design and implement safe deployment patterns, use AI-driven systems to improve signal detection and reduce noise, create and maintain evidence-based maturity assessments
Seniority
Senior, hands-on IC