SRE Engineer (Full Time; Multiple Openings)
Core
Ensuring RingCentral's cloud-based communications products are highly available, reliable, secure, and scalable by managing massive multi-tenant production environments.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Cloud-based unified business communications service (Message Video Phone platform)
Domain
Cloud infrastructure, telecommunications, distributed systems
Deliverable
production ML models | infrastructure
Required skills
Python, Bash, Go, Terraform, Ansible, AWS, GCP, Kubernetes, GitLab, DNS, Docker, CI/CD, TCP/IP, Linux, SRE principles, DevOps practices, system automation, monitoring tools
Preferred skills
Root cause analysis, non-functional requirements definition, cloud technology innovation, automated testing, self-healing systems design
Technologies
Python, Bash, Go, Terraform, Ansible, AWS, GCP, Kubernetes, GitLab, Docker, Linux
Responsibilities
Maintain 24x7 production environment with high service availability; Partner with development teams on production software deployment improvements; Implement automation and orchestration for cloud service operations; Develop internal OPS/SRE tools; Perform quality reviews and manage operational issues; Interface with Dev/QNOPS teams for root cause analysis and prevention of outages; Define non-functional requirements for scalable distributed systems; Implement automated tests, deployments, and proactive monitoring; Collaborate with Product and Support teams on product releases
Seniority
Senior, hands-on IC