Speech Software Engineer
Core
Senior Speech Software Engineer building real-time voice AI infrastructure and optimizing ASR/TTS models for enterprise call centers.
Role type
Senior IC speech software engineer (audio/ML)
Builds
Scalable, low-latency streaming ASR → LLM → TTS pipelines for live customer conversations
Domain
Telecommunications / Real-time Voice AI / Speech Processing
Deliverable
production ML models
Required skills
Golang or Python, distributed systems, low-latency streaming, ASR/TTS tuning, audio fundamentals (codecs, jitter, packet loss), WER/CER/MOS metrics, Kubernetes, cloud platforms (AWS/GCP/Azure)
Preferred skills
noise reduction, echo cancellation, VAD, diarization, forced alignment, event-driven architectures, large-scale dataset analysis
Technologies
Golang, Python, Kubernetes, Docker, AWS, GCP, Azure, Opus, G.711
Responsibilities
Tune and optimize ASR/TTS models for noisy environments; Architect and modernize high-availability voice infrastructure; Build multi-threaded server frameworks for concurrent audio streams; Define and implement speech quality evaluation frameworks; Partner with researchers to productionize speech models.
Seniority
Senior, hands-on IC