CareerPlanSign in

AI Infrastructure Engineer

GB London💼 Full-time🗓 2026-07-10 → 2026-09-26

Core

Build systems to train and serve Fin's next generation of AI products, including training pipelines and inference services for custom models.

Role type

Senior AI Infrastructure Engineer (Model Training & Inference)

Builds

Training pipelines for large transformer/LLM models and low-latency inference services for customer support agents.

Domain

AI Infrastructure / Large Language Models / GPU Computing

Deliverable

production ML models

Required skills

Model training at scale, Model inference at scale, Low-level GPU coding (CUDA/Triton), Distributed training, System optimization, Python, Kubernetes, Cloud platforms (AWS)

Preferred skills

Experience at AI-native companies, Production Python in ML contexts, Open source contributions

Technologies

CUDA, Triton, Kubernetes, AWS, Python, Transformers, LLMs

Responsibilities

Implement and scale training pipelines for large transformer and LLM models; Build and optimize inference services with autoscaling and routing; Work on GPU-level performance tuning; Collaborate with ML scientists to bring cutting-edge methods to production; Mentor and develop other engineers; Raise technical standards and operational excellence.

Seniority

Senior, hands-on IC with mentorship responsibilities

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.