CareerPlanGet AI match score →

Computer Vision Researcher (VLM)

London💼 Full-time🗓 2026-05-30 → 2026-07-31

Core

Researching and developing Large Vision-Language Models (VLMs) to bridge 3D spatial geometry with language, enabling machines to reason about and navigate the physical world.

Role type

Senior Research Scientist (Computer Vision & NLP)

Builds

Unified frameworks for spatial intelligence, agentic reasoning systems for robots/drones, and semantic grounding for 3D maps.

Domain

Geospatial AI, 3D Computer Vision, Multimodal Machine Learning

Deliverable

production ML models

Required skills

3D Geometry (SfM, SLAM, VPS), Transformer architectures, VLMs, Multimodal/Semantic understanding, PyTorch or JAX, large-scale data pipelines

Preferred skills

Gaussian Splatting, NeRFs, Robotics (ROS), Agentic systems, Open-set recognition, Zero-Shot learning

Technologies

PyTorch, JAX, ROS

Responsibilities

Architect semantic grounding between 3D features and language embeddings; Develop algorithms for continuous semantics in 3D maps; Build agentic frameworks for embodied AI; Define benchmarks for spatial common sense in LLMs; Mentor researchers in 3D CV and NLP fusion; Partner with product leads on API delivery.

Seniority

Senior, hands-on IC with mentorship

Rewrite
## Job Title Computer Vision Researcher (VLM) ## Company Overview At Niantic Spatial, we’re building the future of geospatial AI. Powered by a proprietary database of over 30 billion posed images and a groundbreaking third-generation digital map, our mission is to develop spatial intelligence that helps both humans and machines better understand, navigate, and engage with the physical world. Our high-fidelity mapping technology unlocks a new dimension of interaction—laying the foundation for AI to truly comprehend and operate within real-world environments. Join us as we build a living model of the world that people and machines can talk to. ## Responsibilities - Architect Semantic Grounding: Lead research into cross-modal grounding that connects 3D spatial features with language embeddings, enabling the LGM to "understand" object relationships and environmental context. - Scale "Understand" Capabilities: Develop and deploy algorithms for continuous semantics, allowing our 3D maps to evolve and improve their situational awareness as new ground-level and aerial data is ingested. - Agentic Frameworks: Build the "spatial brain" for Embodied AI, enabling robots, Drones and other Machines to move beyond simple navigation to mission-level reasoning. - Multimodal Benchmarking: Define the standards for measuring "spatial common sense" in LLMs, creating evaluations that test a model’s ability to interpret and operate within complex 3D scenes. - Technical Mentorship: Serve as the technical anchor for the London R&D hub, resolving architectural disagreements and mentoring the next generation of researchers in the fusion of 3D CV and NLP. - Collaborative Innovation: Partner with Product leads to ensure the "Understand" API delivers high business value for enterprise customers in robotics, logistics, and field operations. ## Requirements - Education: PhD (or equivalent) in Computer Vision, Machine Learning, or Robotics with a focus on Multimodal/Semantic understanding. - Years of Experience: 4+ years of experience in ML research, with a proven track record of shipping models that bridge 3D Vision and Language. - Technical Depth: Expert knowledge of 3D Geometry (SfM, SLAM, VPS) and Transformer-based architectures (VLMs). - Research Impact: Multiple first-author publications at top-tier venues (CVPR, NeurIPS, ICLR) focusing on VLMs, scene understanding or semantic segmentation. - Implementation Mastery: Ability to write production-quality research code in PyTorch or JAX and manage large-scale data pipelines. - Required In-Office Days: 3 days per week ## Nice to Have - Experience with Gaussian Splatting or NeRFs for semantic scene representation. - Background in robotics (ROS) or building agentic systems that interact with physical environments. - Experience with "open-set" recognition and Zero-Shot learning. ## Benefits - Work in a cutting-edge geospatial AI company. - Collaborate with top researchers and engineers. - Contribute to impactful projects in spatial intelligence and AI. - Opportunity to mentor and lead research initiatives. - Access to advanced tools and infrastructure for research and development. Candidate Privacy Policy I understand that by submitting my job application, the information I provide as part of that application will be used in accordance with Niantic Spatial’s Privacy Notice for Job Applicants and Candidates https://www.nianticspatial.com/applicant-privacy-notice. If required by law, by submitting my job application I consent to the processing of my information as described in that Notice, including processing information I voluntarily disclose to Niantic Spatial, such as health or medical information, race or ethnicity data, and sexual orientation data and, in limited circumstances sharing information with third parties such as references and other third parties that assist in the hiring process. Niantic Spatial is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, citizenship, color, family or medical care leave, gender identity or expression, genetic information, immigration status, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran or military status, race, ethnicity, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance with reasonable accommodation during the application process, please contact your recruiter.
Sourced via ashby · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Ashby ↗