Machine Learning Engineer, AI Labs
Core
Build, optimize, and deploy enterprise-scale AI/ML solutions, specifically focusing on LLM inference optimization and high-performance serving infrastructures for the Netskope Intelligent Security Service Edge (SSE) platform.
Role type
Senior Machine Learning Engineer (LLM Inference & Distributed Systems)
Builds
Production-grade LLM serving infrastructure and AI/ML systems for cloud security
Domain
Cloud Security / Artificial Intelligence / Distributed Systems
Deliverable
production ML models
Required skills
LLM inference optimization, distributed systems architecture, high-throughput system design, memory management (KV Cache), system scalability, production deployment, technical strategy execution
Preferred skills
vLLM, SGLang, TensorRT-LLM, advanced KV Cache optimization techniques
Technologies
vLLM, SGLang, TensorRT-LLM
Responsibilities
Design and optimize enterprise-scale LLM serving infrastructures to improve throughput and latency; Collaborate with senior architects to drive AI/ML technical strategies; Partner with ML scientists to translate business requirements into deployed code; Implement and scale production model performance tracking and reporting systems
Seniority
Senior, hands-on IC