Senior Solutions Engineer, AI Infrastructure
Core
Customer-facing technical role designing, evaluating, and deploying infrastructure for large-scale AI, HPC, analytics, and data-intensive workloads.
Role type
Senior Solutions Architect (AI Infrastructure)
Builds
GPU clusters, high-performance storage systems, Kubernetes platforms, distributed training environments, inference platforms, and data pipelines. (via careerplan.io/jobs/43-001-3E-B68-senior-solutions-engineer-ai-infrastructure-at-vast-data)
Domain
AI Infrastructure, HPC, Cloud, Distributed Systems
Deliverable
production ML models
Required skills
Linux kernel internals, distributed systems (PAXOS, raft), storage implementations, high-performance networking, Kubernetes, GPU infrastructure, MLOps, systems debugging
Preferred skills
Large-scale GPU clusters, petabyte-scale storage, orchestration systems (Slurm, Ray, Spark), distributed filesystems (Lustre, Ceph, Weka), InfiniBand/RoCE/RDMA, CUDA/NCCL
Technologies
Kubernetes, Slurm, Ray, Spark, Lustre, Ceph, Weka, BeeGFS, GPFS, VAST, InfiniBand, RoCE, RDMA, NVIDIA/Mellanox, CUDA, NCCL, DCGM, GPUDirect
