Staff Software Engineer, Machine Learning Platform
Core
Technical lead for the ML Platform team, designing and building infrastructure that enables ML engineers and data scientists to ship models and features reliably at scale.
Role type
Staff Software Engineer (ML Platform)
Builds
End-to-end ML infrastructure, including low-latency model inference, large-scale feature stores, real-time monitoring, and LLM/agent orchestration systems.
Domain
Financial technology / Machine Learning Platform Engineering
Deliverable
production ML models | infrastructure
Required skills
Service-oriented architecture, large-scale distributed systems, technical leadership, system design, cross-functional collaboration, product instincts, ambiguity management, AI tool usage
Preferred skills
Large-scale serving/data infrastructure, LLMs and agentic AI patterns, rapid prototyping, cloud AI/ML services, training and shipping ML models, technical vision setting
Technologies
AWS, SageMaker, Bedrock, Databricks, OpenAI
Responsibilities
Take ownership of end-to-end architecture for complex ML Platform projects; Define technical directions for high-ambiguity projects; Design solutions for challenging problems like low-latency inference and feature stores; Lead large projects from requirements to production; Translate ML team needs into scalable technical solutions; Arbitrate critical decisions balancing latency, reliability, cost, and security; Advise senior leadership on the end-to-end ML lifecycle; Drive cross-team initiatives to improve MLOps maturity; Mentor engineers and model software system design.
Seniority
Staff, hands-on IC with significant mentorship and strategy impact