Sr. Solutions Engineer (GPU Infrastructure)
Core
Primary technical partner for customers deploying and operating GPU clusters and AI infrastructure on Nebius, bridging customer requirements with internal engineering capabilities.
Role type
Senior Solutions Engineer (GPU Infrastructure)
Builds
Full-stack AI cloud platform supporting developers and enterprises from data/model training to production deployment
Domain
Cloud infrastructure, GPU orchestration, AI/ML
Deliverable
client delivery
Required skills
GPU infrastructure expertise, Linux system administration, cluster-level troubleshooting, hardware/software diagnostics, solution architecture, technical documentation
Preferred skills
NVIDIA Grace Blackwell platform experience, cluster validation and benchmarking tools (HPL, NCCL)
Responsibilities
Act as main technical interface for customers running workloads on Nebius GPU infrastructure; Support customers in deploying, configuring, and tuning GPU-based environments; Investigate and resolve complex issues spanning hardware, networking, and OS; Convert customer requirements into practical architectures and execution plans; Develop and maintain technical documentation including solution patterns and troubleshooting guides
Seniority
Senior, hands-on IC
