Engineering Manager, Safeguards Review Tooling
Core
Lead the engineering team building investigation, review, and enforcement tooling for Anthropic's first-party products and third-party cloud platforms to ensure safe model deployment.
Role type
Engineering Manager, Safety Tooling
Builds
Review tooling platform including analytics, privacy-preserving data access primitives, sandbox environments, and Claude-assisted human-in-the-loop workflows
Domain
AI Safety / Trust & Safety / Platform Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
Engineering team management, full-stack or platform engineering, internal tooling delivery, cross-functional partnership with policy/legal/ops, technical architecture design
Preferred skills
Trust and safety tooling experience, privacy/compliance system design, LLM integration in operational workflows, developer platform design, multi-surface enforcement systems
Technologies
Claude, LLMs, analytics platforms, privacy-preserving primitives
Responsibilities
Lead and develop a team of engineers building investigation and enforcement tooling; Define vision and roadmap for review tooling platform; Drive strategy for scaling review through automation; Partner with policy, operations, legal, and data science stakeholders; Ensure tooling evolves alongside privacy commitments; Create clarity in ambiguous environments; Coach top technical talent
Seniority
Senior, hands-on IC with management responsibilities