Senior Machine Learning Engineer, Voice Agents - EMEA Remote
Core
Building the open-source speech-to-speech library and the hf-voice product to enable developers to build and deploy realtime voice agents.
Role type
Senior Machine Learning Engineer (Voice Agents & Infrastructure)
Builds
Open-source speech-to-speech library, hf-voice product API, realtime inference serving infrastructure
Domain
AI/ML, Voice Agents, Developer Tools
Deliverable
production ML models | product features | infrastructure
Required skills
Architectural ownership, Async Python, Distributed systems, Realtime streaming (WebSockets/WebRTC), LLMs/Multimodal models, Developer API design, GPU serving
Preferred skills
Voice-agent frameworks (pipecat, LiveKit, Vocode), Low-level inference runtimes (llama.cpp), ASR/TTS/End-to-end speech models, Audio pipelines (VAD, echo cancellation), Embedded/robotics deployment
Technologies
Python, WebSockets, WebRTC, GPU, Async frameworks
Responsibilities
Design pipeline architecture and latency budgets, Integrate new ASR/TTS models, Build developer API and streaming protocols, Implement realtime inference serving with autoscaling, Write documentation and examples, Support existing deployments (Reachy Mini)
Seniority
Senior, hands-on IC