CareerPlanSign in

大模型评估产品经理(J82456)

北京市💼 Full-time🗓 2026-07-21 → 2026-09-28

Core

Define evaluation standards for large language models (Q&A, creative content, information processing) and analyze results to guide model and system strategy optimization.

Role type

Product Manager for Large Model Evaluation

Builds

Evaluation standards, analysis insights, and automated evaluation tools

Domain

Artificial Intelligence / Large Language Models

Deliverable

production ML models

Required skills

Evaluation standard formulation, deep analysis of evaluation results, product iteration collaboration, competitive analysis, automated evaluation tool design

Preferred skills

User behavior research, business analysis, strong writing sense, pattern recognition, technical sensitivity to deep learning and machine learning, structured thinking

Technologies

Large Language Models, Deep Learning, Machine Learning, NLP, Dialogue Robots

Responsibilities

Establish evaluation standards for LLMs, analyze evaluation results to guide optimization, collaborate with product and engineering teams, conduct market and competitor analysis, design automated evaluation tools

Seniority

Mid-level, hands-on IC

Sourced via baidu · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.