Research Engineer, Multimodal Reasoning For Information Literacy
Core
Research and build multimodal reasoning systems and Vision-Language Models (VLMs) to assess the trustworthiness of media (images, audio, and videos) on the internet.
Role type
Research Engineer, Multimodal Reasoning
Builds
Multimodal models for media authenticity assessment and AI-assisted information literacy tools
Domain
AI safety, online information quality, computer vision, multimodal machine learning
Deliverable
production ML models
Required skills
Computer vision techniques, multimodal machine learning, Python, deep learning frameworks (Jax, TensorFlow, PyTorch), quantitative math and statistics, data analysis and visualization
Preferred skills
Large-scale model training and deployment, video understanding, large language models, prompt engineering, few-shot learning, post-training techniques, peer-reviewed research publications
Technologies
Jax, TensorFlow, PyTorch
Responsibilities
Plan and perform rapid prototyping of computer vision and multimodal ML techniques for media authenticity; Design and train multimodal models for complex visual reasoning; Undertake exploratory analysis to inform research directions; Engage with product teams to drive research development; Implement tools, libraries, and frameworks to accelerate research; Report and present research findings and experimental results; Collaborate with internal and external scientific domain experts
Seniority
Mid-level Research Engineer