Senior Reliability Engineer
Core
Lead enterprise-wide reliability and infrastructure projects, architect scalable cloud solutions, and provide tier 2/3 technical support to enterprise customers for an identity security platform.
Role type
Staff Site Reliability Engineer (Infrastructure & Customer Support)
Builds
Scalable Cloud Prem and SaaS infrastructure solutions for enterprise customers
Domain
Identity Security, Multi-cloud Infrastructure, Compliance
Deliverable
production ML models | infrastructure
Required skills
Cloud infrastructure architecture, Kubernetes, Terraform, Python/Go scripting, incident management, compliance automation, distributed systems, technical leadership, customer troubleshooting
Preferred skills
Bazel, CueLang, compliance-as-code patterns, automated compliance scanning
Technologies
AWS (VPC, EC2, RDS, EKS, CloudFormation), Kubernetes, Helm, Linux, Terraform, GitOps, CI/CD, Prometheus, Grafana, DataDog, Bazel, CueLang
Responsibilities
Lead enterprise-wide reliability projects, architect and deploy cloud solutions at scale, respond to critical incidents and lead root cause analysis, develop automation for operations, provide tier 2/3 technical support to enterprise customers, mentor SRE team members, partner with Security teams on compliance controls
Seniority
Staff, strategic leadership with hands-on technical execution