Site Reliability Engineer
Core
Design, develop, and implement strategies to improve the availability, latency, performance, and efficiency of Netskope's cloud security platform and production services.
Role type
Site Reliability Engineer (SRE)
Builds
Cloud security platform services and microservice architectures
Domain
Cloud Security / Distributed Systems
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Microservices architecture, Cloud services (public/private), System performance tuning, Root cause analysis, Automation, Capacity planning, Non-functional requirements analysis, Code debugging and optimization
Preferred skills
TCP/IP and network flow analysis, IaC and CI/CD tools, Python/C/C++/Go/Rust, Docker/Kubernetes, AWS/GCP, Agile methodologies
Technologies
Python, C, C++, Go, Rust, Docker, Kubernetes, AWS, GCP, KVM, OpenNebula, OpenStack
Responsibilities
Partner with development teams to architect highly available and secure features; Develop innovative ways to monitor and report on service and infrastructure health; Drive efficiencies in systems and processes including capacity planning and performance tuning; Design, develop, test, and implement automation and solutions for production services; Act as a subject matter expert on designated products and infrastructure.
Seniority
Mid-Senior, hands-on IC