AI Infra 实习生(J104435)
Core
Build and optimize distributed training frameworks and inference engines for large-scale foundation and specialized models (text, image, voice) to support business deployment.
Role type
AI Infrastructure Engineer (Intern)
Builds
Distributed training frameworks, inference engines, and reinforcement learning toolchains for trillion-parameter models.
Domain
Artificial Intelligence / Large Language Models / Distributed Systems
Deliverable
production ML models
Required skills
C/C++/Python, PyTorch, Megatron, DeepSpeed, CUDA programming, GPU cluster tuning, vLLM, SGlang, model compression, knowledge distillation, speculative decoding, KV cache optimization
Preferred skills
Agentic workflows, multi-turn environment interaction, offline validation pipelines
Responsibilities
Optimize SFT, RL, and model compression techniques; build efficient distributed training frameworks; optimize inference engines for throughput and cost; construct reinforcement learning toolchains.
Seniority
Intern