视觉多模态理解算法实习生(J96196)
Core
Develop and research multimodal understanding application models to advance state-of-the-art technology and improve application performance.
Role type
Research intern (multimodal understanding)
Builds
Multimodal understanding models
Domain
Computer Vision, Natural Language Processing, Machine Learning
Deliverable
production ML models
Required skills
Multimodal learning, Deep learning, Computer Vision, Natural Language Processing, Algorithm implementation, Code proficiency
Preferred skills
Publications in top-tier conferences (CVPR/ICCV/ECCV/AAAI/ACL/EMNLP/NeurIPS), High ranking in programming or academic competitions (ACM/Kaggle)
Sourced via baidu · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.