Staff Engineer, Machine Learning Platform
Core
Designing and building scalable ML infrastructure and platforms to enable data scientists and ML engineers to ship models and features reliably at low latency.
Role type
Staff Engineer, Machine Learning Platform
Builds
End-to-end ML platform services, model inference systems, feature stores, and orchestration tools for ML-driven products.
Domain
Financial technology / Machine Learning Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Service-oriented architecture, large-scale distributed systems, technical leadership, system design, cross-functional collaboration, product instincts, ambiguity management, AI tool usage
Preferred skills
Large-scale serving/data infrastructure, LLMs and agentic AI patterns, rapid prototyping, cloud AI/ML services, model training and shipping, technical vision, distributed team management
Technologies
AWS, SageMaker, Bedrock, Databricks, OpenAI
Responsibilities
Define long-term strategy and technical direction for next-gen ML infrastructure; design system architecture for low-latency inference and large-scale feature stores; lead large projects from requirements to production operation; translate ML team needs into scalable technical solutions; arbitrate technical decisions balancing latency, reliability, cost, and security; mentor engineers and drive cross-team MLOps initiatives.
Seniority
Staff, strategic technical leadership with hands-on architecture