CareerPlanSign in

AI Infra研发工程师(J99641)

上海市💼 Full-time🗓 2026-07-21 → 2026-09-28

Core

Optimizing training and inference of large language models on domestic GPU hardware and software platforms.

Role type

Senior IC AI Infrastructure Engineer (LLM Optimization)

Builds

Optimized training/inference pipelines and evaluation infrastructure for domestic GPU-based AI models.

Domain

AI Infrastructure / Domestic GPU Computing

Deliverable

production ML models

Required skills

Python, C/C++, PyTorch, Distributed Training, Model Compression, Quantization, GPU Architecture Knowledge, Containerization, Algorithm & Data Structures

Preferred skills

Megatron, vLLM, Zero/Offload, MoE Architecture Tuning, Heterogeneous Computing Acceleration, Large-scale Training/Inference Experience

Technologies

PyTorch, Megatron, vLLM, Domestic GPU Chips

Responsibilities

Optimize LLM training and inference on domestic GPUs; Accelerate software/hardware platforms and operators; Research and implement frontier optimization technologies; Build and maintain evaluation infrastructure for fast/accurate feedback.

Seniority

Mid-Senior, hands-on IC

Sourced via baidu · Listed on CareerPlan, which tracks 845,000+ jobs from 20+ sources.