Member of Technical Staff
Core
Design and implement improvements to training infrastructure, optimize model performance, and work on post-training processes including reinforcement learning and fine-tuning.
Role type
Senior GPU Performance Engineer
Builds
Training infrastructure, model serving infrastructure, and optimized large-scale AI models
Domain
Artificial Intelligence / Deep Learning / GPU Computing
Deliverable
production ML models
Required skills
Python, PyTorch, CUDA, C++, large-scale model training, GPU workload profiling, performance tuning, cluster scaling (Slurm, Kubernetes)
Preferred skills
Experience with reinforcement learning, fine-tuning, low-level GPU code optimization
Technologies
PyTorch, CUDA, C++, Slurm, Kubernetes
Responsibilities
Design and implement improvements to training infrastructure; optimize performance of GPU-accelerated workloads; contribute to model serving infrastructure efficiency; work on post-training processes including RL and fine-tuning.
Seniority
Senior, hands-on IC