CareerPlanSign in

视觉多模态理解算法实习生(J96196)

北京市,上海市,深圳市💼 Full-time🗓 2026-07-21 → 2026-09-28

Core

Develop and research multimodal understanding application models to advance state-of-the-art technology and improve application performance.

Role type

Research intern (multimodal understanding)

Builds

Multimodal understanding models

Domain

Computer Vision, Natural Language Processing, Machine Learning

Deliverable

production ML models

Required skills

Multimodal learning, Deep learning, Computer Vision, Natural Language Processing, Algorithm implementation, Code proficiency

Preferred skills

Publications in top-tier conferences (CVPR/ICCV/ECCV/AAAI/ACL/EMNLP/NeurIPS), High ranking in programming or academic competitions (ACM/Kaggle)

Sourced via baidu · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.