CareerPlanGet AI match score →

AI Infrastructure Engineer

Berlin, Germany💼 Full-time🗓 2026-05-31 → 2026-07-31

Core

Building the systems that train and serve Fin's next generation of AI products, including training pipelines and inference services for custom models.

Role type

Senior AI Infrastructure Engineer (Model Training & Inference)

Builds

Training pipelines for large transformer/LLM models and low-latency inference services for customer support agents.

Domain

AI Infrastructure / Large Language Models / GPU Computing

Deliverable

production ML models

Required skills

Model training at scale, Model inference at scale, Low-level GPU coding (CUDA/Triton), Distributed training, Autoscaling and routing, GPU kernel tuning, Python

Preferred skills

Kubernetes, AWS, AI native company experience

Technologies

CUDA, Triton, Kubernetes, AWS, Python

Responsibilities

Implement and scale training pipelines for large transformer and LLM models; Build and optimize inference services with autoscaling and routing; Work on GPU-level performance tuning; Collaborate with ML scientists to bring cutting-edge methods to production; Mentor and develop other engineers; Raise technical standards and operational excellence.

Seniority

Senior, hands-on IC

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗