大模型评测研发工程师-AI数据与安全
Core
Building engineering infrastructure and AI Agents for large model evaluation, including dataset management, human/machine evaluation capabilities, and automated delivery systems.
Role type
Senior IC machine-learning engineer (evaluation & agent systems)
Builds
Evaluation infrastructure, automated evaluation agents, and high-performance distributed systems for model assessment
Domain
Artificial Intelligence / Large Language Models / Model Evaluation
Deliverable
production ML models
Required skills
Full-stack development, distributed system design, LLM principles, Agent framework development, data structure and algorithms, system design
Preferred skills
LLM training experience, LLM-as-a-judge implementation, multi-agent system design, open-source contributions
Technologies
Agent frameworks, distributed storage, middleware, frontend frameworks
Responsibilities
Develop evaluation infrastructure for dataset ingestion and management; Build AI Agents to enable automated and high-quality model evaluation; Design and implement complex Agent systems for specific business problems; Collaborate with algorithm and product teams to solve complex engineering challenges.