Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI
Core
Building post-training algorithms and optimizing ML systems for next-gen agent RL training platforms to serve enterprise clients.
Role type
Senior IC machine learning systems research engineer (agent post-training)
Builds
Next-gen agent training algorithms, multi-agent/multi-tool rollouts, and optimized training/inference frameworks for enterprise GenAI.
Domain
Generative AI, Large Language Models (LLMs), Reinforcement Learning (RL), Enterprise AI
Deliverable
production ML models
Required skills
LLM training in production, post-training methods (RLHF/RLVR, PPO/GRPO), multi-node LLM training and inference, GPU cluster architecture, CUDA, PyTorch, transformers, flash attention
Preferred skills
System optimization, cross-functional collaboration
Technologies
CUDA, PyTorch, transformers, flash attention
Responsibilities
Build, profile, and optimize training and inference frameworks; Post-train state-of-the-art models to define stable recipes; Collaborate with ML teams to accelerate R&D; Create next-gen agent training algorithms for multi-agent/multi-tool rollouts.
Seniority
Senior, hands-on IC