CareerPlanSign in

Safeguards Enforcement Lead, Cyber Harms

San Francisco, CA💼 Full-time💰 $285,000–$285,000🗓 2026-09-12 → 2026-09-26

Core

Lead enforcement actions to detect and mitigate misuse of AI systems for malicious cyber operations, malware development, and offensive exploitation.

Role type

Senior management IC in AI safety and cyber threat enforcement

Builds

Strategic enforcement frameworks and operational capabilities for AI misuse detection

Domain

AI Safety / Cybersecurity / Threat Intelligence

Deliverable

production ML models | product features | dashboards & analysis | client delivery | infrastructure

Required skills

People management, offensive cybersecurity techniques, malware analysis, vulnerability research, content review at volume, SQL/Python for threat detection, stakeholder communication, generative AI prompt engineering

Preferred skills

Trust & safety experience, LLM misuse understanding, abuse monitoring systems, policy implementation at scale, government agency collaboration

Technologies

SQL, Python, Generative AI models, Abuse monitoring systems

Responsibilities

Manage team of Cyber Enforcement Analysts and contractors, develop strategies to detect AI misuse for cyberattacks, collaborate on high-severity cases, partner with Engineering/Data Science on tooling, stay updated on threat actor tactics

Seniority

Senior, hands-on IC with management responsibilities

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.