Sr. ML Engineer
Core
Design, build, and operate scalable ML platform infrastructure and tooling to support Data Scientists and AI Engineers in moving models from research to production.
Role type
Senior IC ML Platform Engineer
Builds
Scalable ML pipelines, orchestration frameworks, model serving infrastructure, and secure cloud/on-prem environments
Domain
Enterprise AI/ML infrastructure, Generative AI, LLMs, Cloud Computing
Deliverable
production ML models | infrastructure
Required skills
AWS services (EC2, S3, EKS, SageMaker, IAM, VPC, CloudWatch), Kubernetes, Docker, ML pipeline orchestration (Kubeflow, Airflow, MLflow), Infrastructure as Code (Terraform, CloudFormation), CI/CD for ML, Python, Shell scripting, GPU orchestration, ML serving frameworks (vLLM, TensorRT-LLM, KServe, Triton), Distributed computing (Spark)
Preferred skills
Generative AI/LLMOps experience, AI-assisted productivity tools (GitHub Copilot, ChatGPT), Mentoring junior engineers, Hybrid cloud/on-prem architecture, Legacy ML pipeline modernization
Technologies
AWS, Kubernetes, Docker, Terraform, Kubeflow, Airflow, MLflow, vLLM, TensorRT-LLM, KServe, Triton, Spark, Python, Shell
Responsibilities
Lead platform engineering initiatives and guide the team on scalable infrastructure patterns; Build tooling to simplify model deployment and productionization; Design and build scalable ML pipelines and model serving infrastructure; Collaborate with cross-functional teams to bring AI/ML solutions to production; Build and operate secure cloud and on-prem infrastructure; Support GPU-enabled infrastructure for training and inference; Modernize legacy ML pipelines and adopt emerging technologies
Seniority
Senior, hands-on IC with mentorship responsibilities
