多模态算法工程师-抖音内容理解
Core
Building and training foundational multimodal large models (CV, NLP, audio) to enhance content understanding and generation capabilities for Douyin, Toutiao, and Xigua Video.
Role type
Senior IC multimodal algorithm engineer (large language models & generative AI)
Builds
Production multimodal foundation models and training infrastructure for search, recommendation, and advertising scenarios
Domain
Social media, video streaming, and generative AI
Deliverable
production ML models
Required skills
deep learning, large language models, multimodal models, generative models, mathematical foundations, data engineering, model training, training/inference framework iteration
Preferred skills
experience in short video/image algorithms, publications in top-tier conferences (NIPS, ICML, CVPR, etc.), competition experience
Technologies
PyTorch, TensorFlow, distributed training frameworks, model evaluation systems