AI Infrastructure Engineer
Core
Building the systems that train and serve Fin's next generation of AI products, including training pipelines and inference services for custom models.
Role type
Senior AI Infrastructure Engineer (Model Training & Inference)
Builds
Training pipelines for large transformer/LLM models and low-latency inference services for customer support agents.
Domain
AI Infrastructure / Large Language Models / GPU Computing
Deliverable
production ML models
Required skills
Model training at scale, Model inference at scale, Low-level GPU coding (CUDA/Triton), Distributed training, Autoscaling and routing, GPU kernel tuning, Python
Preferred skills
Kubernetes, AWS, AI native company experience
Technologies
CUDA, Triton, Kubernetes, AWS, Python
Responsibilities
Implement and scale training pipelines for large transformer and LLM models; Build and optimize inference services with autoscaling and routing; Work on GPU-level performance tuning; Collaborate with ML scientists to bring cutting-edge methods to production; Mentor and develop other engineers; Raise technical standards and operational excellence.
Seniority
Senior, hands-on IC