MTS, AI Engineering, SMAI
Core
Architect and execute large-scale custom model training and fine-tuning jobs on multi-GPU clusters to power Micron's machine learning and GenAI solutions for manufacturing processes.
Role type
Senior GPU Performance Engineer (AI/ML)
Builds
Scalable AI/ML solutions, autonomous AI Agents, and optimized GPU workloads for manufacturing.
Domain
Semiconductor manufacturing, High-Performance Computing (HPC), Generative AI
Deliverable
production ML models
Required skills
GPU architecture expertise, distributed training strategies, LLM fine-tuning, GenAI application development, high-performance kernel programming, C++, Python
Preferred skills
HPC job schedulers, lower-level optimization, Multi-Agent Systems, computer vision, signal processing
Technologies
CUDA, HIP, PyTorch, vLLM, TensorRT-LLM, LangChain, LangGraph, LlamaIndex, AutoGen, Kubernetes, Docker, Jenkins
Responsibilities
Architect and execute large-scale custom model training and fine-tuning jobs; Optimize training throughput and memory efficiency; Design and develop autonomous AI Agents; Analyze and profile complex workloads; Write and optimize high-performance kernels; Collaborate with Hardware Architects to define GPU features; Design and implement performance regression testing suites; Mentor junior engineers on parallel programming.
Seniority
Senior, hands-on IC