CareerPlanGet AI match score →

Applied Researcher, Audio Understanding

*HQ - San Francisco, CA💼 Full-time🗓 2025-09-16 → 2026-07-31

Core

Architect and develop novel, large-scale models for complex audio understanding tasks, including multi-speaker ASR, diarization, and non-speech audio classification, deploying them to production at scale.

Role type

Senior Applied Researcher (Audio Understanding)

Builds

Large-scale pre-training and fine-tuning datasets for audio understanding capabilities

Domain

AI, Audio Perception, Machine Learning

Deliverable

production ML models

Required skills

ASR, audio understanding, language modeling, generative modeling, large-scale training, GPU/TPU acceleration, model optimization

Preferred skills

self-supervised learning, few-shot learning, audio-visual perception

Technologies

SSMs, GPU, TPU

Responsibilities

Architect and develop novel, large-scale models for complex audio understanding tasks; Pioneer research in self-supervised learning for audio, few-shot learning, and robust audio-visual perception; Set new standards for evaluating and benchmarking audio understanding systems; Build large scale pre-training and fine-tuning datasets for audio understanding capabilities

Seniority

Senior, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Ashby ↗