Research Intern - AI Evaluation and Alignment
Core
Research Intern developing evaluation frameworks and benchmarking methods to assess LLM quality, robustness, and generalization.
Role type
Research Intern (AI Evaluation and Alignment)
Builds
Evaluation frameworks and benchmarking methods for large language models
Domain
Artificial Intelligence / Machine Learning / Large Language Models
Deliverable
production ML models
Required skills
LLM project experience, Python coding, statistical/computational modeling
Preferred skills
Reward modeling, LLM-as-a-Judge, deep learning frameworks, software engineering best practices
Technologies
PyTorch, TensorFlow, Git
Responsibilities
Co-develop research projects with supervisors, design and implement machine learning approaches, develop evaluation frameworks and benchmarking methods, present research findings
Seniority
Intern
Sourced via microsoft · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.