混元大模型评测算法研究员(北京)
Core
Designing and implementing evaluation benchmarks, automated toolchains, and data production pipelines for general-purpose AI large models across text, multimodal understanding, and generation modalities.
Role type
Senior IC research scientist (large model evaluation)
Builds
Automated evaluation toolchains, multimodal evaluation benchmarks, and domain-enhanced evaluation datasets
Domain
Artificial Intelligence / Large Language Models / Multimodal AI
Deliverable
production ML models
Required skills
Machine learning fundamentals, Deep learning model implementation, Data analysis, Logical reasoning, Automated toolchain design, Multimodal data processing
Preferred skills
Large model tuning and optimization, Evaluation experience, First-author publications at top ML conferences (NeurIPS, ICML, KDD, AAAI, IJCAI)
Sourced via tencent · Listed on CareerPlan, which tracks 846,000+ jobs from 20+ sources.