Humanities Evaluation Specialist | Remote , Contract
Core
Develop high-difficulty question-and-answer pairs and test prompts to train next-generation AI models on humanities interpretation, argument, and analytical reasoning.
Role type
Contractor AI evaluation specialist (humanities domain)
Builds
Training data for AI models
Domain
Artificial Intelligence + Humanities
Deliverable
production ML models
Required skills
Humanities studies expertise, research & source triangulation, analytical thinking, written precision, attention to detail, self-direction
Preferred skills
Advanced academic background in humanities, experience producing nuanced questions, familiarity with AI model evaluation
Technologies
None stated
Responsibilities
Develop original, high-difficulty Q&A pairs challenging AI models on interpretation and reasoning; Research, verify, and provide authoritative citations referencing primary texts; Design prompts demanding deep analytical insight; Test questions against AI models and iterate for clarity and correctness; Write precise, unambiguous questions and well-supported answers; Apply feedback to refine submissions and maintain quality standards