CareerPlanSign in

大语言模型强化学习算法专家 - Seed Model

北京💼 Full-time🗓 2026-09-28

Core

Developing and optimizing large language models (LLMs) to push intelligence limits, focusing on scaling laws, system efficiency, and alignment for applications like generation, reasoning, and code.

Role type

Senior IC reinforcement learning algorithm expert (LLM)

Builds

Production LLMs and multi-modal capabilities for consumer apps (e.g., Doubao) and enterprise clients via Volcano Engine.

Domain

Artificial Intelligence / Large Language Models / Reinforcement Learning

Deliverable

production ML models

Required skills

C/C++ or Python, data structures and algorithms, NLP, computer vision, large model training, RL algorithms

Preferred skills

impactful LLM projects or papers, deep system optimization, scaling law research

Technologies

C/C++, Python, NLP, CV

Responsibilities

Discover and apply simple, universal ideas to optimize models across scales; explore boundaries of ultra-large models for performance and efficiency; research next-gen compute scaling directions; lead data construction, instruction tuning, preference alignment, and model optimization; drive application landing for generation, reasoning, and code; research future use cases to expand model scope.

Sourced via bytedance · Listed on CareerPlan, which tracks 853,000+ jobs from 20+ sources.