CareerPlanGet AI match score →

Machine Learning Researcher, Multimodal LLMs

San Francisco🌐 Remote💼 Full-time💰 $180,000–$180,000🗓 2026-04-21 → 2026-07-31

Core

Developing next-generation multimodal LLM stacks that combine speech, text, tools, and real-time reasoning into a single unified system for conversational AI agents.

Role type

Senior IC machine learning researcher (multimodal LLMs)

Builds

Industry-leading conversational AI models powering real-time voice agents

Domain

AI/ML, Voice Technology, Conversational Agents

Deliverable

production ML models

Required skills

LLMs, multimodal models, speech-language systems, prompting, fine-tuning, alignment techniques, neural audio codecs, experimental design, system architecture

Preferred skills

real-time voice systems, tool-using agents, multimodal datasets, open source contributions

Technologies

LLM frameworks, neural audio codecs, agent frameworks

Responsibilities

Design and execute experiments to validate modeling ideas, integrate streaming audio and tool execution into coherent systems, translate abstract modeling concepts into user-facing improvements, take ownership from research through deployment

Seniority

Senior, hands-on IC

Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Ashby ↗