CareerPlanSign in
💼 Full-time🗓 2026-06-25

Core

Building and maintaining the global infrastructure and reliability systems for an AI-powered language learning platform serving millions of users.

Role type

Senior Platform/SRE Engineer (Cloud-Native)

Builds

Scalable, reliable infrastructure on GCP supporting Node.js, Postgres, and Redis services for a global user base.

Domain

EdTech / AI / Cloud Infrastructure

Deliverable

infrastructure

Required skills

Kubernetes, GCP, Terraform, Node.js, Python, PostgreSQL, Redis, Prometheus, Sentry, CI/CD pipeline design, incident management, root cause analysis, SLO/SLA definition, systems thinking, infrastructure automation

Preferred skills

Cost optimization, security, chaos engineering, disaster recovery planning

Technologies

GCP, Kubernetes, Terraform, Node.js, Python, PostgreSQL, Redis, Prometheus, Sentry

Responsibilities

Own infrastructure reliability across GCP and Kubernetes; lead P0/P1 incident response and postmortems; improve observability and on-call processes; define and drive SLO/SLA adoption; build tools for safer deploys and automation; collaborate with Product and ML teams on reliability; set platform roadmaps.

Seniority

Senior, hands-on IC with mentoring responsibilities

Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.