CareerPlanSign in

Speech Software Engineer

New York💼 Full-time🗓 2026-01-12 → 2026-09-29

Core

Senior Speech Software Engineer building real-time voice AI infrastructure and optimizing ASR/TTS models for enterprise call centers.

Role type

Senior IC speech software engineer (audio/ML)

Builds

Scalable, low-latency streaming ASR → LLM → TTS pipelines for live customer conversations

Domain

Telecommunications / Real-time Voice AI / Speech Processing

Deliverable

production ML models

Required skills

Golang or Python, distributed systems, low-latency streaming, ASR/TTS tuning, audio fundamentals (codecs, jitter, packet loss), WER/CER/MOS metrics, Kubernetes, cloud platforms (AWS/GCP/Azure)

Preferred skills

noise reduction, echo cancellation, VAD, diarization, forced alignment, event-driven architectures, large-scale dataset analysis

Technologies

Golang, Python, Kubernetes, Docker, AWS, GCP, Azure, Opus, G.711

Responsibilities

Tune and optimize ASR/TTS models for noisy environments; Architect and modernize high-availability voice infrastructure; Build multi-threaded server frameworks for concurrent audio streams; Define and implement speech quality evaluation frameworks; Partner with researchers to productionize speech models.

Seniority

Senior, hands-on IC

Sourced via lever · Listed on CareerPlan, which tracks 855,000+ jobs from 20+ sources.