大模型算法工程师(语音)-火山方舟
Core
Research and develop next-generation AI core technologies, specifically focusing on audio generation, speech, and multimodal understanding.
Role type
Senior IC large model algorithm engineer (audio/NLP)
Builds
Pre-trained models and downstream applications for audio, NLP, and multimodal tasks
Domain
Artificial Intelligence / Large Language Models / Audio Processing
Deliverable
production ML models
Required skills
Pre-trained model development, Reinforcement learning, Efficient training, Speech synthesis and recognition, Natural language processing, C/C++ or Python, Data structures and algorithms
Preferred skills
Publications in top-tier conferences (NeurIPS, ICML, ICLR, CVPR, ACL, KDD), Awards in programming/AI competitions (ACM, ICPC, Kaggle), Familiarity with specific large models (Seed-TTS, CosyVoice, Qwen-Omni, Veo 3), Independent problem-solving
Technologies
Seed-TTS, CosyVoice, Qwen-Omni, Veo 3
Responsibilities
Research and develop multimodal models, Apply technologies to business scenarios for audio/NLP generation, Investigate and track frontier technologies in audio/NLP/multimodal fields
Seniority
Senior, hands-on IC