北京-语音算法工程师(J100777)
Core
Research and development of algorithms for speech recognition (ASR), speech synthesis (TTS), speaker verification, and speech enhancement, applying large model technologies to speech understanding and generation tasks.
Role type
Machine Learning Engineer (Speech/Audio)
Builds
Production speech models and multimodal fusion applications
Domain
Audio/Speech Technology
Deliverable
production ML models
Required skills
Deep learning, Python, speech signal processing, ASR/TTS models, large language models
Preferred skills
Publications in Interspeech/ICASSP, experience with Whisper/Codec LM
Technologies
Python, Deep Learning Frameworks, Whisper, Codec LM
Sourced via baidu · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.