Safeguards Enforcement Analyst, Safety Evaluations
Core
Enforce safety policies and monitor AI model evaluations to ensure models meet safety standards before and after launch.
Role type
Safeguards Enforcement Analyst (Safety Evaluations)
Builds
Evaluation processes, frameworks, and tooling for AI safety assessments
Domain
AI Safety / Trust and Safety
Deliverable
production ML models
Required skills
program management, process building, cross-functional coordination, data analysis, documentation, risk identification, stakeholder management, ambiguity navigation
Preferred skills
zero-to-one program building, SOP creation, SQL/dashboards proficiency, working with sensitive content
Technologies
SQL, dashboards, spreadsheets, internal AI tools (Claude Code)
Responsibilities
Run and monitor model evaluations to detect regressions or unexpected behavior; coordinate creation of new evaluations and update existing ones; interpret evaluation results and drive mitigations; build processes and frameworks for product-specific evaluations; design tooling improvements for self-serve eval creation; write and maintain rigorous documentation for evaluation workflows
Seniority
Individual Contributor, mid-to-senior level