Senior Staff Site Reliability Engineer
Core
Define and own the SRE strategy for LiveRamp's global infrastructure, setting technical direction for reliability engineering across the organization.
Role type
Senior Staff Site Reliability Engineer (Strategy & Architecture)
Builds
Global data collaboration platform infrastructure supporting consumer privacy and identity use cases
Domain
Cloud Infrastructure, Data Collaboration, Consumer Privacy
Deliverable
production ML models | infrastructure
Required skills
SRE strategy definition, Infrastructure as Code (Terraform), Kubernetes internals and autoscaling, distributed systems architecture, Python/Go, observability engineering (SLOs/SLIs), FinOps, CI/CD platform maturity, cloud security (IAM, network segmentation)
Preferred skills
Multi-region active-active architectures, chaos engineering, LLMs/AI-assisted development workflows
Technologies
Terraform, Kubernetes, SingleStore, ScyllaDB, Cassandra, DynamoDB, Python, Go, Jenkins, CircleCI, GCP, AWS
Responsibilities
Define and own SRE strategy including SLOs/SLAs and error budgets; oversee automation of critical areas to mitigate risk; develop complex software infrastructure spanning multiple products; lead distributed systems architecture reviews; serve as escalation point for high-impact production incidents; champion FinOps strategy across cloud resources; mentor Staff Engineers; establish production readiness standards
Seniority
Senior Staff, hands-on IC with organization-wide scope