CareerPlanGet AI match score →

Safeguards Enforcement Analyst, Violence & Extremism

San Francisco, CA💼 Full-time💰 $285,000–$285,000🗓 2026-07-15 → 2026-07-31

Core

Build and execute operational workflows to assess AI model behavior, drive enforcement decisions, and develop evals to mitigate misuse for violence, extremism, and dangerous technology.

Role type

Safeguards Enforcement Analyst (Violence & Extremism)

Builds

Automated enforcement systems, review workflows, and evals for AI safety

Domain

AI Safety / Counterterrorism / Threat Intelligence

Deliverable

production ML models | product features

Required skills

Policy enforcement, threat intelligence, SQL/data analysis, generative AI product experience, risk identification, stakeholder communication

Preferred skills

Weapons/dangerous tech expertise, legal/regulatory frameworks, red-teaming AI, threat actor profiling, OSINT, LLM technical understanding, Python

Technologies

SQL, Python, MITRE ATT&CK, LLMs

Responsibilities

Design automated enforcement systems and review workflows; Develop evals to measure model performance and surface regressions; Partner with Engineering/Data Science to optimize detection; Review flagged content to drive enforcement decisions; Support policy design with feedback on gaps; Develop enforcement guidelines and documentation; Track emerging misuse patterns and extremist activity

Seniority

Mid-Senior, hands-on IC

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗