Sr Site Reliability Engineer
Core
Ensuring reliability, observability, and operational excellence for a high-traffic real estate platform serving millions of users.
Role type
Senior Site Reliability Engineer (IC)
Builds
Highly available AWS infrastructure (EKS, Fargate), CI/CD pipelines (Skyway, CircleCI, Argo CD), and API gateways (Tyk, Apollo).
Domain
Real Estate / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
AWS (EKS, Fargate, Lambda, VPC), Kubernetes, Python/Go/Java, Terraform, CloudFormation, NewRelic, Splunk, CI/CD (CircleCI, Argo CD), Incident Response, Chaos Engineering.
Preferred skills
Tyk/Kong, Apollo GraphQL, FinOps, Service Mesh (Istio), Datadog, Grafana, Prometheus, Vault, OpsGenie, PagerDuty.
Responsibilities
Implement and maintain multi-region AWS infrastructure; monitor SLIs/SLOs and error budgets; execute chaos engineering experiments; build dashboards and alerts for rapid troubleshooting; optimize infrastructure costs via FinOps; participate in on-call rotation and post-incident reviews.
Seniority
Senior, hands-on IC