CareerPlanGet AI match score →

Research Engineer, Code RL (Reinforcement Learning)

San Francisco, CA💼 Full-time💰 $500,000–$500,000🗓 2026-06-11 → 2026-07-31

Core

Design RL environments and coding tasks, build reward signals and verifiers for code quality, run training experiments on frontier models, and improve pipelines for autonomous software engineering.

Role type

Research Engineer (Reinforcement Learning for Code Generation)

Builds

RL systems for agentic coding behaviors, code correctness, long-horizon autonomous engineering, and high-performance code for accelerators.

Domain

Artificial Intelligence / Reinforcement Learning / Software Engineering

Deliverable

production ML models

Required skills

Python (async/concurrent programming), software engineering, system design, experimental design, result interpretation, code quality assurance, performance optimization

Preferred skills

Reinforcement Learning, RLHF, LLM finetuning, program analysis, testing, verification, compilers, formal methods, PyTorch, large-scale distributed training, CUDA/GPU/TPU kernel experience, virtualization, sandboxed code execution

Technologies

Python, PyTorch, CUDA, GPU, TPU

Responsibilities

Design RL environments and coding tasks; build reward signals and verifiers; run training experiments on frontier models; diagnose model performance in software engineering tasks; improve pipeline speed and reliability.

Seniority

Mid-Senior, hands-on IC

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗