大模型策略产品经理-安全方向(J92070)
Core
Design and execute human and automated evaluation frameworks for large language model (LLM) safety, collaborating with algorithm teams to optimize RL training strategies and enhance model performance in weak areas.
Role type
Senior LLM Strategy Product Manager (Safety)
Builds
Automated evaluation benchmarks and optimized RL training environments for LLMs
Domain
Artificial Intelligence / Large Language Models / Safety
Deliverable
production ML models
Required skills
LLM safety evaluation design, Reinforcement Learning strategy, Python data analysis, System architecture design, AI tool orchestration
Preferred skills
Experience with complex scenario benchmarking, Data-driven decision making
Technologies
Python, RL training environments, AI tools
Responsibilities
Build automated evaluation benchmarks covering complex scenarios; Analyze model weaknesses and design targeted RL training strategies; Collect and analyze data to support benchmark optimization and training adjustments; Track industry trends and competitor dynamics; Apply AI tools to optimize workflow and data processing.
Seniority
Senior, hands-on IC