Senior Platform Engineer - W/M/NB
Core
Design, build, and operate a GPU platform on GCP for AI workloads, enabling self-service compute and model serving for Data Scientists and Machine Learning Engineers.
Role type
Senior Platform Engineer (MLOps/DevOps)
Builds
GPU infrastructure-as-code, self-service Kubernetes clusters (Ray), and golden paths for service deployment.
Domain
Cloud Infrastructure (GCP) + AI/ML Platform Engineering
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, Terraform, GitOps, Distributed Systems, GCP, Software Engineering fundamentals
Preferred skills
GPU workload optimization, Distributed job scheduling (KubeRay/Slurm/Kubeflow), Cross-cloud networking, MLOps tooling (MLflow/W&B)
Technologies
GCP, Kubernetes, Terraform, ArgoCD, Helm, Ray
Responsibilities
Design and operate GPU platform on GCP; Deliver self-service compute with quotas and cost visibility; Establish golden paths for service deployment; Oversee infrastructure costs; Automate internal workflows.
Seniority
Senior, hands-on IC