混元大模型音频理解算法工程师(北京)
Core
Develop large-scale audio understanding models including ASR, audio captioning, and multimodal audio-video understanding.
Role type
Senior IC machine-learning engineer (audio understanding)
Builds
Production audio understanding models and open-source model releases
Domain
Artificial Intelligence / Audio Processing / Large Language Models
Deliverable
production ML models
Required skills
Python, PyTorch, Megatron, FSDP, SLIME, VERL, data structures, algorithms
Preferred skills
ASR, audio understanding, music understanding, multimodal learning, pre-training, fine-tuning, reinforcement learning, ACM/ICPC/NOI/IOI/Top Coder/Kaggle awards, top-tier conference publications (NeurIPS/ICLR/ICML/ACL/CVPR/ICASSP/INTERSPEECH)
Sourced via tencent · Listed on CareerPlan, which tracks 846,000+ jobs from 20+ sources.