Senior DevOps Engineer – AI Platform / Kubernetes / AWS / GPU Infrastructure - CDI
Core
Design, secure, and scale cloud and on-prem infrastructure supporting Mistral AI and Prisme AI platforms for generative AI workloads in a critical banking environment.
Role type
Senior DevOps / Platform Engineer (AI Infrastructure)
Builds
Enterprise-scale LLM orchestration platforms, GPU-intensive AI workloads, and MLOps/LLMOps pipelines.
Domain
Banking / Generative AI / Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Kubernetes administration, AWS hybrid cloud, CI/CD pipeline construction, Infrastructure as Code, GPU resource allocation, observability stack implementation, security standards (IAM, RBAC, encryption), MLOps/LLMOps concepts.
Preferred skills
Generative AI platform experience, self-hosted LLM deployment, NVIDIA stack and GPU operators, high-performance inference serving, SRE practices, regulated industry background.
Technologies
AWS, Kubernetes, Docker, Helm, Kustomize, GitLab CI, GitHub Actions, ArgoCD, Terraform, Ansible, Prometheus, Grafana, ELK, Loki, OpenTelemetry, Mistral AI, Prisme AI.
Responsibilities
Design and maintain highly available cloud and on-prem infrastructure for generative AI platforms; deploy and administer Kubernetes clusters for AI/LLM workloads; optimize CPU/GPU/memory/storage allocation and scaling; build and industrialize CI/CD pipelines for AI models and APIs; implement advanced observability stacks for AI metrics; enforce enterprise security standards and regulatory compliance.
Seniority
Senior, hands-on IC