MLOps / Platform Engineer (m/w/d)
Core
Building and operating the infrastructure behind AI systems, focusing on Kubernetes clusters and platform services.
Role type
MLOps/Platform Engineer
Builds
Central cluster services, CI/CD and GitOps pipelines, observability stacks for production systems
Domain
AI/ML infrastructure, Cloud-native platforms
Deliverable
infrastructure
Required skills
Kubernetes, CI/CD pipelines, GitOps, Linux, Networking, Observability (logging, metrics, alerting), Software Engineering, DevOps
Preferred skills
Ray, GPU infrastructure, ArgoCD, ArgoWorkflows, Prometheus, Grafana, Loki
Responsibilities
Ensure reliable operation and configuration of central cluster services, improve platform services (Ingress, certificates, DNS, secrets), maintain CI/CD and GitOps pipelines, integrate new services into clusters, monitor production services and build out observability
Seniority
Mid-level to Senior, hands-on IC
