Research Engineer (LLM Performance), London
Core
Optimize large-scale LLM training and inference systems to accelerate drug discovery and molecular design.
Role type
Senior Research Engineer (LLM Performance)
Builds
Production-ready systems for supervised fine-tuning, reinforcement learning, and LLM evaluation
Domain
Biopharmaceuticals / Large Language Models / Distributed Systems
Deliverable
production ML models
Required skills
Large-scale distributed LLM training, Deep learning frameworks (JAX/PyTorch), Parallelism strategies, Collective communication libraries (NCCL), GPU architecture optimization
Preferred skills
LLM serving stacks, Accelerator DSLs (XLA, Triton, Pallas, CUDA), Low-precision optimization, GCP production systems
Technologies
JAX, PyTorch, NCCL, GCP, CUDA, XLA, Triton, Pallas
Responsibilities
Implement and optimize LLM post-training methods at scale, Collaborate with research teams to translate methods into production systems, Diagnose and fix performance bottlenecks in distributed training and inference, Deploy low-precision methods to balance performance and accuracy
Seniority
Senior, hands-on IC