AI Engineer (Voice & Speech)
Core
Develop advanced Speech AI and Voice AI capabilities at the intersection of research, product engineering, and production.
Role type
Senior IC AI Engineer (Speech & Voice)
Builds
End-to-end audio AI pipelines, voice assistants, and speech-enabled product features
Domain
Artificial Intelligence / Speech Processing / Audio Machine Learning
Deliverable
production ML models
Required skills
Python, Deep Learning (PyTorch/TensorFlow), Speech AI (ASR/TTS/Audio AI), Model Deployment, End-to-End ML Workflows, Data Engineering, Performance Optimization
Preferred skills
Foundation Models (Whisper/wav2vec/NVIDIA NeMo), Mobile/Edge Deployment, Real-time Audio Streaming, Speaker Recognition/Diarization, Voice Cloning, LLM Integration
Technologies
PyTorch, TensorFlow, Whisper, wav2vec 2.0, NVIDIA NeMo, XTTS, VITS, SpeechBrain, Kaldi
Responsibilities
Research and prototype new speech AI approaches; Develop and optimize intelligent voice technologies (ASR, TTS, audio processing); Design and optimize production-ready audio AI pipelines; Collaborate with product teams to build voice assistants and conversational experiences; Deploy and monitor AI models in production environments; Build and improve high-quality datasets for model training and evaluation
Seniority
Senior, hands-on IC