AI Infrastructure Engineer
Core
Build systems to train and serve Fin's next generation of AI products, including training pipelines and inference services for custom models.
Role type
Senior AI Infrastructure Engineer (Model Training & Inference)
Builds
Training pipelines for large transformer/LLM models and low-latency inference services for customer support agents.
Domain
AI Infrastructure / Large Language Models / GPU Computing
Deliverable
production ML models
Required skills
Model training at scale, Model inference at scale, Low-level GPU coding (CUDA/Triton), Distributed training, System optimization, Python, Kubernetes, Cloud platforms (AWS)
Preferred skills
Experience at AI-native companies, Production Python in ML contexts, Open source contributions
Technologies
CUDA, Triton, Kubernetes, AWS, Python, Transformers, LLMs
Responsibilities
Implement and scale training pipelines for large transformer and LLM models; Build and optimize inference services with autoscaling and routing; Work on GPU-level performance tuning; Collaborate with ML scientists to bring cutting-edge methods to production; Mentor and develop other engineers; Raise technical standards and operational excellence.
Seniority
Senior, hands-on IC with mentorship responsibilities