高性能计算专家(深圳/北京)
Core
Build and optimize high-performance computing platforms for AI training and inference workloads, focusing on GPU/NPU hardware, drivers, and compute libraries.
Role type
Senior IC high-performance computing expert (GPU/NPU)
Builds
Industry-leading high-performance computing platforms
Domain
Artificial Intelligence / High-Performance Computing
Deliverable
production ML models
Required skills
GPU hardware architecture expertise, CUDA/ROCm programming, multi-GPU/NPU optimization, deep learning framework optimization (PyTorch/TensorFlow), compute library development, kernel virtualization, performance profiling and debugging
Preferred skills
Academic research experience, publications, industry foresight
Technologies
NVIDIA CUDA, AMD ROCm, PyTorch, TensorFlow, NVIDIA GPUs, Ascend NPUs, Intel/AMD processors
Responsibilities
Optimize underlying performance of compute resources (NVIDIA/Ascend/Intel/AMD), troubleshoot hardware/driver/compute library issues for multi-card systems, conduct academic research and publish papers, evaluate and integrate emerging technologies into existing systems
Seniority
Senior, hands-on IC