Research Scientist- Vision-Language-Action (VLA) Models
Core
Conduct research and engineering in core AI and machine learning fields to enable Embodied AI (including computer vision, autonomous planning, open-world learning) for ADAS/AD, industrial automation, and robotics.
Role type
Research Scientist (Vision-Language-Action Models)
Builds
Modular end-to-end perception and planning systems for ADAS/AD, industrial automation, and robotics
Domain
Automotive (ADAS/AD), Robotics, Industrial Automation
Deliverable
production ML models
Required skills
Computer Vision, Robotic Motion and Behavioral Planning, Reinforcement Learning, Multimodal Transformers, Deep Learning, Python, C++, Rust, TensorFlow, PyTorch
Preferred skills
Real-world product development of autonomous systems, Multimodal vision-language-action models, Diffusion models, NeRF, gaussian splatting, 3D scene understanding, sensor calibration, SfM, voxel/BEV grid-based feature representation
Responsibilities
Conduct research and engineering in core AI and machine learning fields to enable Embodied AI, Push boundaries in modular end-to-end perception and planning for ADAS/AD, Collaborate cross-functionally with global research and engineering teams, Implement research results to solve real-world challenges, Engage with academic and industry communities through conferences and workshops, Document and disseminate research findings through publications and patents
Seniority
Senior, hands-on IC