Senior Software Engineer (Voice Backend)
Core
Design, build, and scale low-latency backend systems and real-time infrastructure for voice AI agents, including media pipelines and streaming orchestration.
Role type
Senior IC backend engineer (voice/real-time systems)
Builds
Real-time voice backend infrastructure, low-latency audio streaming pipelines, and orchestration layers for ASR/LLM/TTS integration.
Domain
Enterprise AI, Voice Technology, Distributed Systems
Deliverable
production ML models
Required skills
Distributed systems engineering, Real-time audio streaming (WebRTC, SIP, RTP), ML inference serving (ASR/TTS), High-concurrency service design, Dialogue state tracking, Event-driven architectures (Pub-Sub), Cloud infrastructure (Kubernetes, AWS/GCP/Azure), Database management (PostgreSQL, ClickHouse)
Preferred skills
Telephony platforms (Twilio, LiveKit, Pipecat, Asterisk/FreeSWITCH), A/B testing frameworks, Python, Golang, Java
Technologies
WebRTC, SIP, RTP, WebSocket, gRPC, Docker, Kubernetes, PostgreSQL, ClickHouse, AWS, GCP, Azure, Python, Golang, Java
Responsibilities
Design and build real-time voice backend with low-latency audio streaming pipelines and orchestration layers; Manage session and conversation-state for thousands of concurrent calls; Build high-throughput services for serving speech and language models; Integrate and operate streaming ASR/TTS engines; Engineer reliability features like failover and observability; Design APIs and data plane connecting voice runtime to the platform; Lead technical initiatives to integrate voice backend and partner with ML engineers.
Seniority
Senior, hands-on IC