No visa sponsorship at Artificial Analysis
Core
Evaluate frontier AI systems (language, image, video, speech, hardware) across quality, speed, and pricing to inform AI labs, enterprises, and policymakers.
Role type
Senior AI/ML Engineer (Applied AI Research & Evaluation)
Builds
AI evaluation benchmarks, inference systems, and analysis reports for frontier model organizations
Domain
Artificial Intelligence / Machine Learning / Benchmarking
Deliverable
production ML models
Required skills
Applied AI research, language-model evaluations, inference optimization, speech processing, AI hardware knowledge, robotics
Preferred skills
Experience with frontier models, evaluation framework design
Technologies
Python, PyTorch, Hugging Face, LLMs, GPU clusters
Responsibilities
Design and execute evaluations for frontier language and multimodal models, analyze inference performance and costs, collaborate with AI labs to define evaluation metrics, publish industry analysis
Seniority
Senior, hands-on IC