GPU资源运营(AI方向) - 火山引擎
Core
Manage and optimize GPU cluster resources for AI workloads, balancing performance and cost to meet internal and external client computing demands.
Role type
GPU resource operations engineer (AI infrastructure)
Builds
GPU cloud services, AI model training/inference platforms, and custom resource solutions for enterprise clients.
Domain
Cloud computing and AI infrastructure
Deliverable
infrastructure
Required skills
GPU cluster management, AI server hardware knowledge, resource planning, utilization optimization, public cloud GPU services understanding, cross-functional collaboration
Preferred skills
AI training/inference scenario expertise, video rendering experience, SRE collaboration
Technologies
GPU clusters, AI servers, public cloud GPU services
Responsibilities
Introduce and operate GPU clusters, plan and allocate resources, optimize utilization, evaluate hardware configurations, deliver custom resource solutions, collaborate on operational tool development.
Seniority
Mid-level, hands-on IC