Research Engineer / Scientist, Post-training - Paris
Core
Develop and train advanced Large Language Models (LLMs) and Vision-Language Models (VLMs) to optimize capabilities for agentic AI applications, focusing on instruction following, tool use, and reinforcement learning.
Role type
Senior Research Engineer / Scientist (Post-training)
Builds
Foundational LLMs and VLMs for agentic AI systems
Domain
Artificial Intelligence, Machine Learning, Deep Learning
Deliverable
production ML models
Required skills
Python, Git, PyTorch, JAX, TensorFlow, large-scale distributed training, LLM training, alignment, reinforcement learning, multimodal architectures
Preferred skills
Publications in top-tier AI conferences (NeurIPS, ICML, CVPR, ACL, ICCV), industry experience, data processing techniques
Technologies
PyTorch, JAX, TensorFlow, Git
Responsibilities
Develop and train advanced LLMs and VLMs including multimodal architectures; Research and implement training methods for enhanced capabilities like instruction following and tool use; Design and optimize data pipelines and training systems for large-scale distributed training; Collaborate with cross-functional teams to integrate models into agentic AI systems; Evaluate model performance and communicate findings to stakeholders; Stay current with advancements in LLMs, VLMs, and related fields
Seniority
Senior, hands-on IC