CareerPlanGet AI match score →

Product Manager, Safeguards (Verticals)

San Francisco, CA💼 Full-time💰 $305,000–$305,000🗓 2026-07-15 → 2026-07-31

Core

Designing and deploying safety systems, evaluations, and interventions to protect Anthropic's frontier AI models and products from misuse and risks across cloud platforms.

Role type

Product Manager, AI Safety & Safeguards

Builds

Safeguards systems, detection tools, evaluation frameworks, and product UX for AI safety

Domain

Artificial Intelligence / AI Safety / Generative AI

Deliverable

production ML models | product features | dashboards & analysis

Required skills

Product strategy, safety system design, technical tradeoff analysis, stakeholder alignment, risk mitigation planning, metrics development, ambiguous environment navigation, zero-to-one product building, complex technical communication

Preferred skills

5+ years in product management, experience with AI/ML research teams, data detection and intervention expertise, infrastructure and tools knowledge, evaluation (evals) experience

Technologies

AI safety tools, detection systems, evaluation frameworks, cloud platforms

Responsibilities

Define upstream safety-by-design strategies and downstream defenses for frontier models; write safety evaluations and communicate safety externally; prioritize solutions and define requirements for MVP vs. ideal state; collaborate with policy, research, and engineering teams; plan mitigation for deployment risks of powerful models; develop metrics to measure performance and blindspots

Seniority

Mid-to-Senior, hands-on IC

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗