CareerPlanSign in

AI Engineer, Platforms

NTU Main Campus, Singapore💼 Full-time🗓 2026-07-29 → 2026-09-26

Core

Design and implement optimized inference workflows and support building customized LLM-based solutions for production deployment.

Role type

Senior IC AI Platform Engineer (LLM Infrastructure)

Builds

Multi-cloud LLM inference platforms, API services, and GPU-enabled AI infrastructure

Domain

Artificial Intelligence / Cloud Infrastructure / Large Language Models

Deliverable

production ML models

Required skills

LLM inference optimization, GPU cluster management, Infrastructure as Code, CI/CD pipeline development, Observability engineering, REST API design, Model Context Protocol (MCP), Python scripting, Cloud provider operations

Preferred skills

Multimodal AI model experience, C/C++/Rust/Go programming, Open-source AI/ML contributions

Technologies

vLLM, SGLang, TensorRT-LLM, Terraform, Docker, Kubernetes, Claude, Copilot, Cursor

Responsibilities

Operate and monitor multi-cloud LLM inference platform (SEA-LION API Farm), optimize LLM inference across modalities, build new API services, manage high-performance AI clusters and storage systems, develop CI/CD pipelines and deployment automation, strengthen observability across the stack

Seniority

Mid-level, hands-on IC

Sourced via workday · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.