Research Program Manager - Model Evals and Safety
Core
Build foundational infrastructure, evaluation frameworks, and operational processes for model safety and evaluation from 0 to 1.
Role type
Senior Research Program Manager (Model Evals and Safety)
Builds
Model evaluation frameworks, safety operations workflows, and integration with model development lifecycle
Domain
AI Safety, Model Evaluation, Open-Weight AI
Deliverable
production ML models | infrastructure
Required skills
Technical program management, building new functions from scratch, model evaluation landscape knowledge, red-teaming understanding, stakeholder management, scoping and prioritization of technical investments, external ecosystem engagement
Preferred skills
ML engineering background, experience with pre/mid/post-training pipelines, familiarity with regulatory safety ecosystems
Technologies
N/A
Responsibilities
Define evaluation frameworks and tooling requirements, establish safety operations workflows and review cadences, partner with research/engineering leads to embed safety checkpoints, drive scoping of eval science and infrastructure investments, establish external safety ecosystem engagement, create visibility and reporting structures for leadership
Seniority
Senior, hands-on IC with program leadership