多模态世界模型算法工程师/专家-豆包大模型
Core
Researching and developing multimodal world models, large-scale foundation models, and multimodal agents for virtual and real-world environments to support applications like Doubao and virtual worlds.
Role type
Senior IC multimodal world model algorithm engineer/researcher
Builds
Multimodal understanding and generative foundation models, multimodal agents for GUI/games, and simulation-based environment models
Domain
AI research, multimodal learning, computer vision, AIGC, reinforcement learning
Deliverable
production ML models
Required skills
Multimodal understanding, generative modeling, reinforcement learning, computer vision, large-scale model optimization, data synthesis, scalable oversight, model reasoning, planning
Preferred skills
Top-tier conference publications (CVPR, NeurIPS, etc.), C/C++ or Python proficiency, competitive programming awards, leading high-impact projects in world models or rendering
Technologies
LLM, GenMedia, AIGC, RAG, GUI, simulation environments
Responsibilities
Explore multimodal understanding, generative, and reinforcement learning technologies; Optimize large-scale multimodal foundation models and system performance; Develop multimodal agents for virtual worlds; Model virtual/real-world environments using pre-training and simulation to enable AI-driven applications
