Senior Site Reliability Engineer (SRE) – CloudVision as a Service (CVaaS)
Core
Build and operate a global Kubernetes-native SaaS platform for enterprise network management and streaming telemetry, ensuring scalability, reliability, and stability.
Role type
Senior Site Reliability Engineer (SRE)
Builds
CloudVision as a Service (CVaaS) fleet
Domain
Cloud networking, SaaS, Kubernetes
Deliverable
production ML models | product features | infrastructure
Required skills
Distributed systems architecture, Kubernetes, Python, Golang, Bash, CI/CD, Observability, Capacity planning, Disaster recovery, Cost optimization
Preferred skills
GCP, GKE, Database management
Technologies
Kubernetes, GCP, GKE, Golang, Python, Ansible, Pulumi, Bash
Responsibilities
Drive architecture and performance for the Data Platform (NetDL), lead sustainable incident response and blameless postmortems, implement automation for operational processes.
Seniority
Senior, hands-on IC
