Worldwide Specialist Solutions Architect - AI, Data & AI GTM
Core
Technical authority on model customization and inference to help customers across the AMERICAS unlock the full potential of foundation models, custom training, and production-scale serving on AWS.
Role type
Senior IC Machine Learning Solutions Architect (Model Customization & Inference)
Builds
End-to-end ML architectures spanning data preparation, distributed training, model fine-tuning (LoRA, PEFT, RLHF), inference optimization, and production deployment using SageMaker AI and SageMaker HyperPod.
Domain
Cloud Computing / Generative AI / Machine Learning
Deliverable
production ML models
Required skills
Model customization (fine-tuning, continued pre-training, distillation), Inference optimization (model compilation, quantization, endpoint auto-scaling), Distributed training (FSDP, multi-GPU parallelism), MLOps, Cloud infrastructure design, Solution architecture
Preferred skills
Experience working with end user or developer communities, Community building
Technologies
Amazon SageMaker AI, SageMaker HyperPod, Amazon Bedrock, Trainium, Inferentia, GPU-based EC2, EKS, ECS, LoRA, PEFT, RLHF, FSDP, P5e instances
Responsibilities
Partner with customers' data science and engineering teams to architect scalable ML solutions; Serve as the go-to SME on model customization and inference patterns; Author technical blogs, whitepapers, and reference architectures; Act as the technical liaison between customers and AWS service teams to drive platform improvements; Develop and scale an internal community of ML subject matter experts.
Seniority
Senior, hands-on IC