Anthropic Fellows Program, AI Safety
Core
Empirical AI safety research project producing public outputs (e.g., papers) using external infrastructure.
Role type
Research Fellow (AI Safety)
Builds
Public research outputs (papers, reports) on AI safety topics
Domain
Artificial Intelligence Safety / Machine Learning
Deliverable
production ML models | research
Required skills
Python programming, empirical ML research, large language models
Preferred skills
Open-source contributions, specific AI safety research areas (scalable oversight, adversarial robustness, model internals, AI welfare)
Technologies
Python, large language models
Responsibilities
Conduct empirical research projects, produce public outputs, collaborate with mentors
Seniority
Fellow (early career researcher)
Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.