Product Manager, Safeguards (Verticals)
Core
Designing and deploying safety systems, evaluations, and interventions to protect Anthropic's frontier AI models and products from misuse and risks across cloud platforms.
Role type
Product Manager, AI Safety & Safeguards
Builds
Safeguards systems, detection tools, evaluation frameworks, and product UX for AI safety
Domain
Artificial Intelligence / AI Safety / Generative AI
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Product strategy, safety system design, technical tradeoff analysis, stakeholder alignment, risk mitigation planning, metrics development, ambiguous environment navigation, zero-to-one product building, complex technical communication
Preferred skills
5+ years in product management, experience with AI/ML research teams, data detection and intervention expertise, infrastructure and tools knowledge, evaluation (evals) experience
Technologies
AI safety tools, detection systems, evaluation frameworks, cloud platforms
Responsibilities
Define upstream safety-by-design strategies and downstream defenses for frontier models; write safety evaluations and communicate safety externally; prioritize solutions and define requirements for MVP vs. ideal state; collaborate with policy, research, and engineering teams; plan mitigation for deployment risks of powerful models; develop metrics to measure performance and blindspots
Seniority
Mid-to-Senior, hands-on IC