AI Engineer, Dev Ops
Core
Operate and maintain the central data platform and cloud infrastructure (GCP/AWS) for AI research and products, ensuring reliability, security, and efficiency while integrating AI tools into operational workflows.
Role type
Senior DevOps/SRE Engineer (AI Platform)
Builds
Reliable data pipelines, cloud infrastructure, CI/CD systems, and AI-assisted operational tooling for the Data Platform team.
Domain
AI Research & Development, Cloud Infrastructure, Data Engineering
Deliverable
production ML models | infrastructure
Required skills
Cloud infrastructure operations (GCP, AWS), Infrastructure as Code (Terraform), Container orchestration (Kubernetes, Docker), CI/CD pipeline management, Observability (logs, metrics, traces), Incident response and triage, Data engineering fundamentals, Security (IAM, secrets management, vulnerability patching), Scripting (Python, Bash, Go), AI tool usage for automation and analysis.
Preferred skills
Experience with multimodal data systems, Open-source AI/data tooling contributions, Singapore public sector or research programme experience.
Technologies
GCP, AWS, Terraform, Docker, Kubernetes, Python, Bash, Go, CI/CD tools, Observability stacks, AI tools (Claude, Copilot, Cursor).
Responsibilities
Own day-to-day operations of the Data Platform across GCP and AWS including environment health, capacity, performance, and security posture; Lead incident response including triage, mitigation, and post-mortems; Operate central data management infrastructure including pipelines, storage, and access controls; Manage IaC and maintain CI/CD pipelines and deployment automation; Strengthen observability and automate repetitive operational tasks; Use AI tools deliberately for incident triage, log analysis, runbook drafting, and IaC generation.
Seniority
Mid-Senior, hands-on IC