Engineering Manager, Safeguards
Core
Lead the Review Tooling team to build systems for safety investigators and Claude to investigate harms and enforce actions across Anthropic's products and third-party platforms.
Role type
Engineering Manager, Safety Review Tooling
Builds
Investigation, review, and enforcement tooling; analytics capabilities; privacy-preserving data access primitives; sandbox environments for workflow iteration; automation systems integrating Claude.
Domain
AI Safety / Trust & Safety / Platform Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Engineering team management, full-stack or platform engineering, internal tooling delivery, cross-functional partnership, technical architecture design
Preferred skills
Trust and safety tooling experience, privacy/compliance system design, LLM integration in operational workflows, developer platform design, enforcement system support
Technologies
LLMs, Claude, analytics platforms, privacy-preserving primitives
Responsibilities
Lead and develop engineers building review tooling; define vision and roadmap for the review platform; drive strategy for scaling review via automation; partner with policy, legal, and data science teams; ensure tooling evolves with privacy commitments; create clarity in ambiguous environments; coach top technical talent.
Seniority
Manager, hands-on leadership