星星海实验室-AI Infra 研发工程师
Core
Evaluate and optimize performance of AI acceleration hardware across single-node/single-GPU and super-node scenarios, ensuring software-hardware system tuning and business delivery.
Role type
Senior IC AI infrastructure engineer (hardware performance optimization)
Builds
Optimized AI inference and training systems for enterprise customers
Domain
AI infrastructure, high-performance computing, GPU acceleration
Deliverable
production ML models
Required skills
Linux development, Python, C++, GPU architecture knowledge, distributed training frameworks (Megatron), high-performance network optimization, inference frameworks (TensorRT, vLLM, SGLang)
Preferred skills
Experience with Nvidia/AMD GPU development and optimization
Technologies
Linux, Python, C++, Megatron, TensorRT, vLLM, SGLang, Nvidia GPUs, AMD GPUs
Responsibilities
Conduct comprehensive software-hardware performance evaluation and tuning for AI acceleration hardware; Optimize product line quality and user experience; Track, locate, and resolve technical issues in collaboration with the team.