多模态世界模型算法工程师/专家-豆包大模型
Core
Researching and building multimodal world models, generative AI, and large-scale foundation models to enable virtual/real world interaction and agent capabilities.
Role type
Senior IC multimodal world model algorithm engineer/researcher
Builds
Multimodal understanding and generative foundation models, virtual world agents (GUI/games), and AI-driven new products
Domain
AI research, multimodal learning, generative media, robotics
Deliverable
production ML models
Required skills
Multimodal understanding, generative AI, reinforcement learning, computer vision, large-scale model optimization, data synthesis, scalable oversight, model reasoning, planning
Preferred skills
Top-tier conference publications (CVPR, NeurIPS, etc.), C/C++ or Python proficiency, competitive programming awards, leading high-impact projects in world models or rendering
Technologies
LLM, GenMedia, AIGC, RAG, Visual CoT, GUI, Simulation
Responsibilities
Explore multimodal understanding and generative technologies; Optimize large-scale multimodal foundation models; Build multimodal agents for virtual environments; Model virtual/real world environments using pretraining and simulation
