CareerPlanSign in

北京-语音算法工程师(J100777)

北京市💼 Full-time🗓 2026-07-21 → 2026-09-28

Core

Research and development of algorithms for speech recognition (ASR), speech synthesis (TTS), speaker verification, and speech enhancement, applying large model technologies to speech understanding and generation tasks.

Role type

Machine Learning Engineer (Speech/Audio)

Builds

Production speech models and multimodal fusion applications

Domain

Audio/Speech Technology

Deliverable

production ML models

Required skills

Deep learning, Python, speech signal processing, ASR/TTS models, large language models

Preferred skills

Publications in Interspeech/ICASSP, experience with Whisper/Codec LM

Technologies

Python, Deep Learning Frameworks, Whisper, Codec LM

Sourced via baidu · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.