Senior Research Scientist- Vision-Language-Action (VLA) Models
Core
Conduct research and engineering in core AI and machine learning fields to enable Embodied AI (including computer vision, autonomous planning, open-world learning) for ADAS/AD, industrial automation, and robotics.
Role type
Senior Research Scientist (Vision-Language-Action Models)
Builds
Scalable, intelligent AIoT solutions for automated driving, advanced driver assistance systems (ADAS), robotics, and automation.
Domain
Automotive (ADAS/AD), Robotics, Industrial Automation
Deliverable
production ML models
Required skills
Computer Vision, Robotic/Automotive Motion and Behavioral Planning, Reinforcement Learning (PPO, DQN, DDPG), Multimodal Transformers, Vision-Language-Action Models, Deep Learning, 3D Scene Understanding, Sensor Calibration, SfM, Voxel/BEV Grid-based Feature Representation
Preferred skills
Real-world product development and deployment of autonomous systems, Multimodal language models, Diffusion models, NeRF, Gaussian splatting, Object detection/segmentation
Technologies
Python, C++, Rust, TensorFlow, Pytorch
Responsibilities
Conduct research and engineering in core AI/ML fields; Push boundaries in end-to-end perception and planning; Collaborate cross-functionally for technology transfer and system integration; Implement research results to solve real-world challenges; Engage with academic and industry communities; Document and disseminate research findings through publications and patents.
Seniority
Senior, hands-on IC