Staff Machine Learning Engineer, ML Infrastructure
Core
Building, deploying, and operating cloud-scale ML infrastructure for real-time computer vision inference and LLM/GenAI serving to power intelligent home security products.
Role type
Staff Machine Learning Engineer (ML Infrastructure)
Builds
Kubernetes-based ML platform, real-time CV inference systems, and LLM/GenAI serving infrastructure
Domain
Home Security / Cloud ML Infrastructure
Deliverable
production ML models
Required skills
Kubernetes, Ray, Python, AWS (EKS, S3, IAM), Kafka, CI/CD, Infrastructure-as-Code, GPU-aware scheduling, autoscaling, multi-tenancy, model lifecycle management, incident response, SLO definition
Preferred skills
KServe, Triton, vLLM, Go/C++/Rust, streaming pipelines (Flink), CV workloads, MLflow, open source contributions
Technologies
Kubernetes, Ray, KServe, Triton, vLLM, AWS, Kafka, Python, Go, C++, Rust
Responsibilities
Drive architecture decisions for the ML platform; Build and operate real-time CV inference at scale; Stand up LLM/GenAI serving infrastructure; Mentor engineers and establish best practices; Own reliability and operational excellence
Seniority
Staff, hands-on IC with technical leadership