Performance Engineering Architect
Core
Design, develop, and optimize the performance of AI solutions across GPUs, servers, storage, and networking for HPE's Green Lake cloud platform.
Role type
Senior IC Performance Engineering Architect (AI/GenAI)
Builds
Next-generation cloud platform (Green Lake) and AI workloads for enterprise customers
Domain
Cloud infrastructure, Generative AI, GPU-accelerated systems, Server hardware
Deliverable
Production ML models | product features | dashboards & analysis
Required skills
Generative AI performance optimization, GPU benchmarking, LLM inference frameworks, Linux administration, Python scripting, Server architecture analysis, Performance regression analysis, Telemetry correlation, Test methodology design
Preferred skills
Virtualization platforms (VMware, KVM), Industry benchmarks (SPEC, TPC), Database workloads
Technologies
NVIDIA NIM, vLLM, TensorRT-LLM, PyTorch, Kubernetes, Docker, NVIDIA DCGM, fio, SPEC CPU
Responsibilities
Characterize inference and training performance across models and configurations; Develop custom benchmark suites and automation; Identify performance bottlenecks and recommend hardware/software changes; Establish repeatable test methodologies and release-performance gates; Collaborate with engineering, product, and field teams on solution sizing.
Seniority
Senior, hands-on IC