大语言模型算法研究员(数据智能)-火山引擎
Core
Optimize data synthesis algorithms for pre-training, SFT, and RLHF stages of large language models; support R&D in Code-LLM and logical reasoning models; drive application deployment for RAG-QA robots, data insight robots, and AICoding.
Role type
Senior LLM Algorithm Researcher (Data Intelligence)
Builds
Pre-trained and fine-tuned large language models, NL2Code systems, complex reasoning chains, and deployed AI applications.
Domain
Artificial Intelligence / Large Language Models / Data Intelligence
Deliverable
production ML models
Required skills
Deep learning training frameworks (PyTorch, Huggingface), LLM architecture and training methods, data synthesis algorithms, evaluation systems, technical problem solving
Preferred skills
Top-tier conference papers (ACL, NeurIPS, ICML, EMNLP, ICLR), competition awards, experience with Code-LLM, logical reasoning models, RAG-QA, data insight robots, AICoding, Scaling Law, Post-Training, patent generation
Technologies
PyTorch, Huggingface, Code-LLM, RAG-QA, AICoding
Responsibilities
Optimize data synthesis algorithms for pre-training, SFT, and RLHF; participate in training innovation for Code-LLM and logical reasoning models; explore and iterate on real-world applications like RAG-QA robots and AICoding; follow open-source SOTA models and practice Post-Training in data intelligence.
Seniority
Senior, hands-on IC
