CareerPlanSign in

Member of Technical Staff - RL Environments

London🌐 Remote💼 Full-time🗓 2026-08-17 → 2026-09-26

Core

Building and maintaining reinforcement learning (RL) environments to train and evaluate AI agents for enterprise work.

Role type

Senior IC machine-learning engineer (RL environments)

Builds

RL environments, agent training pipelines, evaluation harnesses, and verifier tools

Domain

Enterprise AI, Reinforcement Learning, Agentic Systems

Deliverable

production ML models

Required skills

engineering AI agents, optimizing agents for industry use cases, reviewing agent trajectories, designing verifier implementations, tuning reward designs, designing annotation workflows, building synthetic data pipelines

Preferred skills

training with RL (scaling, troubleshooting, tuning environments)

Technologies

RL frameworks, agent harnesses, verifier systems, synthetic data generation tools

Responsibilities

Build new RL environments targeting different agentic capabilities and industry areas; Train and evaluate agents in those environments; Make all the pieces work together: tasks, data, tool implementations, and verifiers; Work across modeling and product to identify gaps in agent performance; Work with external vendors to create high-quality, expert-built RL environments; Automate the discovery of model capability gaps and systematically measure agent performance

Seniority

Senior, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.