Applied Researcher, Audio Understanding
Core
Architect and develop novel, large-scale models for complex audio understanding tasks, including multi-speaker ASR, diarization, and non-speech audio classification, deploying them to production at scale.
Role type
Senior Applied Researcher (Audio Understanding)
Builds
Large-scale pre-training and fine-tuning datasets for audio understanding capabilities
Domain
AI, Audio Perception, Machine Learning
Deliverable
production ML models
Required skills
ASR, audio understanding, language modeling, generative modeling, large-scale training, GPU/TPU acceleration, model optimization
Preferred skills
self-supervised learning, few-shot learning, audio-visual perception
Technologies
SSMs, GPU, TPU
Responsibilities
Architect and develop novel, large-scale models for complex audio understanding tasks; Pioneer research in self-supervised learning for audio, few-shot learning, and robust audio-visual perception; Set new standards for evaluating and benchmarking audio understanding systems; Build large scale pre-training and fine-tuning datasets for audio understanding capabilities
Seniority
Senior, hands-on IC