CareerPlanSign in

混元大模型音频理解算法工程师(北京)

Shenzhen, China💼 Full-time🗓 2026-09-28

Core

Develop large-scale audio understanding models including ASR, audio captioning, and multimodal audio-video understanding.

Role type

Senior IC machine-learning engineer (audio understanding)

Builds

Production audio understanding models and open-source model releases

Domain

Artificial Intelligence / Audio Processing / Large Language Models

Deliverable

production ML models

Required skills

Python, PyTorch, Megatron, FSDP, SLIME, VERL, data structures, algorithms

Preferred skills

ASR, audio understanding, music understanding, multimodal learning, pre-training, fine-tuning, reinforcement learning, ACM/ICPC/NOI/IOI/Top Coder/Kaggle awards, top-tier conference publications (NeurIPS/ICLR/ICML/ACL/CVPR/ICASSP/INTERSPEECH)

Sourced via tencent · Listed on CareerPlan, which tracks 846,000+ jobs from 20+ sources.