Research Scientist
Core
Design datasets, evaluations, and RL environments to address gaps in frontier vision-language, multi-modal, robotics, and coding models; run failure analyses on partner labs' models; own research behind new benchmarks; and validate data quality empirically.
Role type
Senior Research Scientist (Applied AI/Benchmarks)
Builds
Open models (RF-DETR), enterprise-grade benchmarks (RF100-VL), and training data for frontier labs
Domain
Artificial Intelligence, Computer Vision, Multi-modal Learning, Robotics
Deliverable
production ML models | research | dashboards & analysis
Required skills
PhD or equivalent in computer vision/multi-modal learning/ML, post-training expertise (fine-tuning, RL, evaluation design), strong engineering fundamentals for experiment code, empirical data validation, benchmark design
Preferred skills
Frontier lab experience, vision-language model expertise, robotics learning, agentic coding evaluations
Technologies
RF-DETR, RF100-VL, RL environments
Responsibilities
Design datasets and RL environments targeting measurable gaps in frontier models; run structured failure analyses on partner labs' models; own research behind new benchmarks; validate data quality through fine-tuning and ablation; contribute to open models; collaborate with frontier lab researchers
Seniority
Senior, hands-on IC with research leadership