CareerPlanSign in

Machine Learning Intern

San Francisco💼 Internship🗓 2026-08-28 → 2026-09-26

Core

Conduct end-to-end research on voice stack components including speech-to-text, large language models, neural audio codecs, and text-to-speech to produce results worth shipping or publishing.

Role type

Research Intern (Audio/ML)

Builds

Production-ready voice AI systems handling millions of calls

Domain

Telephony, Speech AI, Audio Processing

Deliverable

research

Required skills

PyTorch, self-supervised modeling, generative modeling, multimodal modeling, experimental design, ablation studies, GPU cluster management

Preferred skills

Prior publications in speech/language AI, open source contributions, experience with ASR robustness, neural audio codecs, expressive TTS, real-time inference

Technologies

PyTorch, distributed GPU infrastructure

Responsibilities

Design and execute experiments from literature review to implementation, present findings and defend methodology, train/evaluate models on large-scale real-world telephony audio, collaborate with engineers to move results toward production

Seniority

Intern, research-focused

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.