Senior AI Platform Engineer
Core
Architecting and deploying secure, scalable cloud platforms optimized for AI and ML workloads, partnering with AI teams to translate compute needs into infrastructure requirements.
Role type
Senior IC AI Platform Engineer
Builds
Secure, scalable cloud platforms and CI/CD pipelines for AI/ML workloads
Domain
Financial services / Cloud Infrastructure / AI/ML
Deliverable
infrastructure
Required skills
System design, Application development, Testing, Operational stability, Kubernetes, Containerization, Cloud delivery models (IaaS/PaaS/SaaS), Infrastructure as Code, Microservices architecture, AI-assisted development tool governance, Responsible AI workflows
Preferred skills
NVIDIA GPU infrastructure software (DCGM, BCM, Dynamo), Observability tools (Prometheus, Grafana), MLOps tooling (MLflow), High-performance computing frameworks (vLLM, Ray.io, Slurm), Network architecture, Database programming (SQL/NoSQL), Cloud data services, Linux environments
Technologies
Kubernetes, Docker, Python, Go, Java, C#, Prometheus, Grafana, MLflow, vLLM, Ray.io, Slurm, IaaS, PaaS, SaaS, CI/CD, Linux, SQL, NoSQL
Responsibilities
Develop secure, high-quality production code and review/debug code written by others. Architect and deploy secure, scalable cloud platforms optimized for AI and ML workloads. Partner with AI teams to translate compute needs into infrastructure requirements. Monitor, manage, and optimize cloud resources for performance and cost efficiency. Build CI/CD pipelines, automation, and infrastructure-as-code to streamline ML deployment and operations. Drive adoption and governance of approved AI-assisted engineering practices across teams.
Seniority
Senior, hands-on IC