大语言模型强化学习算法专家 - Seed Model
Core
Developing and optimizing large language models (LLMs) to push intelligence limits, focusing on scaling laws, system efficiency, and alignment for applications like generation, reasoning, and code.
Role type
Senior IC reinforcement learning algorithm expert (LLM)
Builds
Production LLMs and multi-modal capabilities for consumer apps (e.g., Doubao) and enterprise clients via Volcano Engine.
Domain
Artificial Intelligence / Large Language Models / Reinforcement Learning
Deliverable
production ML models
Required skills
C/C++ or Python, data structures and algorithms, NLP, computer vision, large model training, RL algorithms
Preferred skills
impactful LLM projects or papers, deep system optimization, scaling law research
Technologies
C/C++, Python, NLP, CV
Responsibilities
Discover and apply simple, universal ideas to optimize models across scales; explore boundaries of ultra-large models for performance and efficiency; research next-gen compute scaling directions; lead data construction, instruction tuning, preference alignment, and model optimization; drive application landing for generation, reasoning, and code; research future use cases to expand model scope.