Platforms Engineer, MLOps
Core
Build and operate modern infrastructure and systems to run large-scale machine learning and deep learning workloads, enabling AI practitioners to develop, train, and deploy models at scale.
Role type
Senior IC MLOps Platforms Engineer
Builds
Production ML infrastructure, tooling stacks, and CI/CD pipelines for ML models
Domain
Artificial Intelligence / Machine Learning Operations / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Python, Go, Rust, or JavaScript/TypeScript, Linux administration, automation tools (Ansible, Terraform, Bash), container orchestration (Docker, Kubernetes, Helm), cloud platforms (AWS, Azure, GCP), infrastructure-as-code, GitOps, observability (Prometheus, Grafana)
Preferred skills
ML experiment tracking (MLflow, Weights & Biases), mentoring technical teams
Technologies
Python, Go, Rust, JavaScript/TypeScript, Linux, Ansible, Terraform, Bash, Docker, Kubernetes, Helm, AWS, Azure, Google Cloud Platform, MLflow, Weights & Biases, Prometheus, Grafana
Responsibilities
Design, build, and maintain platform and tooling stacks; implement MLOps practices including CI/CD pipelines, automated testing, and model monitoring; collaborate with data scientists and engineers to align ML solutions with operational goals; build and maintain resilient, secure, high-performing production infrastructure; document and investigate system issues; develop tools to automate infrastructure; propose and drive technical decisions; mentor apprentices and contribute to MLOps curriculum
Seniority
Senior, hands-on IC