Director, Research - Frontier Benchmarks
Core
Leading a team of researchers to design datasets and benchmarks that define and advance the frontier of AI, specifically focusing on RL and domain-specific evaluations.
Role type
Director, Research - Frontier Benchmarks
Builds
Frontier AI benchmarks, RL training datasets, and evaluation frameworks for model performance.
Domain
Artificial Intelligence / Machine Learning / Research
Deliverable
production ML models
Required skills
Applied AI/ML leadership, benchmark design, data-centric research, roadmap planning, cross-functional collaboration, customer-facing technical translation, team mentorship
Preferred skills
Ph.D. in ML/NLP, experience with agentic evaluation, experience in fast-paced ambiguous environments
Technologies
RL training, agentic evaluation frameworks
Responsibilities
Define data build processes based on market signals and GTM strategy; Own quality standards for benchmark design; Mentor technical teams; Foster external and academic collaborations; Track new research trends in benchmarks and RL training.
Seniority
Director, strategy & mentorship