Specialist Platform Engineering Kubernetes / GPU (m/w/d)
Core
Building and operating a self-service, highly automated, and secure PaaS/CaaS platform with a specific focus on GPU infrastructure for data-intensive applications and deep learning scenarios.
Role type
Platform Engineering Specialist (Kubernetes/GPU)
Builds
Container platforms (Kubernetes, virtualization layers), GPU infrastructure, APIs, CLI tools, and service catalogs for development teams.
Domain
Cloud Infrastructure / Platform Engineering / GPU Computing
Deliverable
production ML models | infrastructure
Required skills
Kubernetes administration and integration, Platform Engineering (CaaS/IaaS) with security and compliance, Infrastructure as Code (IaC) and GitOps, API and CLI tool development, Disaster Recovery and Lifecycle Management, GPU resource planning and virtualization concepts.
Preferred skills
CI/CD pipeline design (GitHub Actions), Cloud environment operations (AWS/Azure/GCP), Monitoring and logging (Prometheus, Grafana, ELK/Loki).
Responsibilities
Build and maintain container platforms with security controls; design and operate GPU infrastructure for AI/ML workloads; implement automation and scaling processes; develop APIs and tools for self-service; provide Level-1 support and consult on architecture and onboarding.
Seniority
Mid-Senior, hands-on IC