CareerPlanSign in

高性能计算专家(深圳/北京)

Beijing, China💼 Full-time🗓 2026-09-28

Core

Build and optimize high-performance computing platforms for AI training and inference workloads, focusing on GPU/NPU hardware, drivers, and compute libraries.

Role type

Senior IC high-performance computing expert (GPU/NPU)

Builds

Industry-leading high-performance computing platforms

Domain

Artificial Intelligence / High-Performance Computing

Deliverable

production ML models

Required skills

GPU hardware architecture expertise, CUDA/ROCm programming, multi-GPU/NPU optimization, deep learning framework optimization (PyTorch/TensorFlow), compute library development, kernel virtualization, performance profiling and debugging

Preferred skills

Academic research experience, publications, industry foresight

Technologies

NVIDIA CUDA, AMD ROCm, PyTorch, TensorFlow, NVIDIA GPUs, Ascend NPUs, Intel/AMD processors

Responsibilities

Optimize underlying performance of compute resources (NVIDIA/Ascend/Intel/AMD), troubleshoot hardware/driver/compute library issues for multi-card systems, conduct academic research and publish papers, evaluate and integrate emerging technologies into existing systems

Seniority

Senior, hands-on IC

Sourced via tencent · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.