感知语义理解算法工程师(J94931)
Core
Develop end-to-end vision-language large models, traffic sign/light recognition, 2D/3D object detection, and scene semantic understanding algorithms for fully autonomous systems.
Role type
Senior IC perception algorithm engineer (vision-language models)
Builds
Perception and semantic understanding models for autonomous driving/robotics
Domain
Autonomous driving / Computer Vision / Robotics
Deliverable
production ML models
Required skills
C++, Python, Linux development, deep learning (3D/convolutional/sparse/transformer), mathematical analysis
Preferred skills
LLM/VLM/VLA training experience, publications in CVPR/ICCV/ICRA/IROS/PAMI
Technologies
C++, Python, Linux, 3D convolution, sparse convolution, transformer
Responsibilities
Research and develop vision-language end-to-end large models, traffic sign/light recognition, 2D/3D object detection, and scene semantic understanding algorithms; propose data requirements and build automated data pipelines; collaborate on fully autonomous perception model development
Seniority
Senior, hands-on IC