CareerPlanGet AI match score →

Engineering Manager, Safeguards Review Tooling

San Francisco, CA💼 Full-time💰 $405,000–$405,000🗓 2026-07-14 → 2026-07-31

Core

Lead the engineering team building investigation, review, and enforcement tooling for Anthropic's first-party products and third-party cloud platforms to ensure safe model deployment.

Role type

Engineering Manager, Safety Tooling

Builds

Review tooling platform including analytics, privacy-preserving data access primitives, sandbox environments, and Claude-assisted human-in-the-loop workflows

Domain

AI Safety / Trust & Safety / Platform Engineering

Deliverable

production ML models | product features | infrastructure

Required skills

Engineering team management, full-stack or platform engineering, internal tooling delivery, cross-functional partnership with policy/legal/ops, technical architecture design

Preferred skills

Trust and safety tooling experience, privacy/compliance system design, LLM integration in operational workflows, developer platform design, multi-surface enforcement systems

Technologies

Claude, LLMs, analytics platforms, privacy-preserving primitives

Responsibilities

Lead and develop a team of engineers building investigation and enforcement tooling; Define vision and roadmap for review tooling platform; Drive strategy for scaling review through automation; Partner with policy, operations, legal, and data science stakeholders; Ensure tooling evolves alongside privacy commitments; Create clarity in ambiguous environments; Coach top technical talent

Seniority

Senior, hands-on IC with management responsibilities

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗