CareerPlanGet AI match score →

Research Engineer, Multimodal Reasoning For Information Literacy

Mountain View, California, US💼 Full-time💰 $174,000–$252,000🗓 2026-05-07 → 2026-07-31

Core

Research and build multimodal reasoning systems and Vision-Language Models (VLMs) to assess the trustworthiness of media (images, audio, and videos) on the internet.

Role type

Research Engineer, Multimodal Reasoning

Builds

Multimodal models for media authenticity assessment and AI-assisted information literacy tools

Domain

AI safety, online information quality, computer vision, multimodal machine learning

Deliverable

production ML models

Required skills

Computer vision techniques, multimodal machine learning, Python, deep learning frameworks (Jax, TensorFlow, PyTorch), quantitative math and statistics, data analysis and visualization

Preferred skills

Large-scale model training and deployment, video understanding, large language models, prompt engineering, few-shot learning, post-training techniques, peer-reviewed research publications

Technologies

Jax, TensorFlow, PyTorch

Responsibilities

Plan and perform rapid prototyping of computer vision and multimodal ML techniques for media authenticity; Design and train multimodal models for complex visual reasoning; Undertake exploratory analysis to inform research directions; Engage with product teams to drive research development; Implement tools, libraries, and frameworks to accelerate research; Report and present research findings and experimental results; Collaborate with internal and external scientific domain experts

Seniority

Mid-level Research Engineer

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗