Site Reliability Engineer
Core
Ensure cloud platform security, reliability, scalability, and operational excellence for IoT smart home products and networking services.
Role type
Site Reliability Engineer (IC)
Builds
Cloud-based microservices on Multi-Cloud Platform (AWS, OCI, Azure, GCP)
Domain
Networking, IoT, Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Kubernetes, Cloud Operations, DevOps, Python, Go, Bash, Chaos Engineering, Load Testing, Incident Response, Technical Documentation, Security Compliance (ISO27001, SOC2, GDPR), KPI Definition (SLA/SLO/SLI)
Preferred skills
AWS/Azure/GCP Solutions Architect Certifications, Container Orchestration
Technologies
Kubernetes, AWS, OCI, Azure, GCP, Python, Go, Bash, Java, PowerShell
Responsibilities
Implement and operate microservices on Kubernetes; Conduct load and chaos tests; Build observability; Write automation scripts; Define and maintain KPIs; Participate in incident response and post-incident analysis; Create technical documentation; Ensure security and compliance adherence; Mentor less senior members; Participate in on-call rotation.
Seniority
Mid-level, hands-on IC