Machine Learning Researcher, Multimodal LLMs
Core
Developing next-generation multimodal LLM stacks that combine speech, text, tools, and real-time reasoning into a single unified system for conversational AI agents.
Role type
Senior IC machine learning researcher (multimodal LLMs)
Builds
Industry-leading conversational AI models powering real-time voice agents
Domain
AI/ML, Voice Technology, Conversational Agents
Deliverable
production ML models
Required skills
LLMs, multimodal models, speech-language systems, prompting, fine-tuning, alignment techniques, neural audio codecs, experimental design, system architecture
Preferred skills
real-time voice systems, tool-using agents, multimodal datasets, open source contributions
Technologies
LLM frameworks, neural audio codecs, agent frameworks
Responsibilities
Design and execute experiments to validate modeling ideas, integrate streaming audio and tool execution into coherent systems, translate abstract modeling concepts into user-facing improvements, take ownership from research through deployment
Seniority
Senior, hands-on IC