Senior AI Engineer, Voice Platform
Core
Own and evolve AI systems for real-time voice platform features including streaming transcription, intelligent reformatting, context-aware mention detection, and voice-to-action pipelines within a productivity tool.
Role type
Senior IC machine-learning engineer (voice platform)
Builds
Real-time speech-to-text pipelines, LLM-powered post-processing, and voice-to-action systems for MAX Desktop, Mobile, Web, and Browser Extension
Domain
Productivity software / Voice AI / Real-time streaming
Deliverable
production ML models
Required skills
Real-time streaming ASR, VAD, audio processing, context injection for transcription, LLM post-processing, natural language parsing, ASR model evaluation and benchmarking, multimodal AI integration
Preferred skills
Whisper, AssemblyAI, Fireworks integration, screen+voice+text multimodal capabilities
Technologies
Whisper, AssemblyAI, Fireworks
Responsibilities
Design and optimize real-time speech-to-text pipelines; Improve transcription accuracy via context injection; Develop LLM-powered post-processing; Build voice-to-action systems; Evaluate and integrate ASR models; Collaborate on shipping voice features across platforms; Explore multimodal AI capabilities
Seniority
Senior, hands-on IC