Machine Learning Intern
Core
Conduct end-to-end research on voice stack components including speech-to-text, large language models, neural audio codecs, and text-to-speech to produce results worth shipping or publishing.
Role type
Research Intern (Audio/ML)
Builds
Production-ready voice AI systems handling millions of calls
Domain
Telephony, Speech AI, Audio Processing
Deliverable
research
Required skills
PyTorch, self-supervised modeling, generative modeling, multimodal modeling, experimental design, ablation studies, GPU cluster management
Preferred skills
Prior publications in speech/language AI, open source contributions, experience with ASR robustness, neural audio codecs, expressive TTS, real-time inference
Technologies
PyTorch, distributed GPU infrastructure
Responsibilities
Design and execute experiments from literature review to implementation, present findings and defend methodology, train/evaluate models on large-scale real-world telephony audio, collaborate with engineers to move results toward production
Seniority
Intern, research-focused