Research Scientist (Speech AI)
Core
Design and implement state-of-the-art machine learning models for speech synthesis and recognition, optimizing performance for real-time inference at scale.
Role type
Senior IC machine learning engineer (speech AI)
Builds
Production speech AI models for voice synthesis and recognition
Domain
Speech AI, deep learning, audio signal processing
Deliverable
production ML models
Required skills
Python, PyTorch or TensorFlow, deep learning architectures (Transformers, CNNs, RNNs), large-scale distributed training, model optimization
Preferred skills
Published research in top-tier ML conferences, experience with speech synthesis models (Tacotron, FastSpeech, VITS), audio signal processing and acoustic modeling, model quantization
Technologies
PyTorch, TensorFlow, Transformers, CNNs, RNNs
Responsibilities
Design and implement ML models for speech synthesis and recognition; Optimize model performance for real-time inference; Build and maintain training pipelines for large-scale model training; Collaborate with research team to implement latest advancements; Work with product team to translate requirements into technical solutions
Seniority
Senior, hands-on IC