CareerPlanGet AI match score →
Onsite or remote • New York City+1💼 Full-time🗓 2026-06-25

Core

Designing, building, and owning complex, scalable, and reliable cloud infrastructure systems end-to-end.

Role type

Senior Cloud/Backend Infrastructure Engineer

Builds

Service infrastructure, internal developer tooling, and scalable data systems

Domain

Cloud Infrastructure & Backend Systems

Deliverable

infrastructure

Required skills

Kubernetes, Cloud platforms (GCP/AWS/Azure), Database scaling (SQL/NoSQL), Streaming systems, Message queues, Observability (DataDog/Grafana), CI/CD pipelines, System architecture design

Preferred skills

Knative, ClickHouse, Pulumi, Temporal

Technologies

Kubernetes, Knative, GCP, Temporal, Pulumi, ClickHouse, MongoDB, Redis, Postgres, BigTable, Firestore, Node.js, TypeScript, Python

Responsibilities

Deploy and support service infrastructure in Kubernetes, Scale and optimize multiple databases, Build internal tooling to accelerate developer velocity, Provide observability and monitoring across systems, Participate in weekend on-call rotation for production incidents

Seniority

Senior, hands-on IC

Rewrite
## About the Role We're looking for a Senior Cloud Backend Engineer to join our Infrastructure Team. You should enjoy solving complex problems, working deep in the details, helping other developers, and owning systems end-to-end. This role also includes being a key part of our weekend on-call rotation, supporting the reliability of our platform. Are you passionate about building reliable, scalable systems? Do you thrive on open-ended problems? Do you love designing, building, and owning complex infrastructure from conception to production? If yes, you're in the right place. ## Responsibilities - Deploy and support our service infrastructure in Kubernetes - Identify the right tools and technologies for major initiatives and then build them - Help other teams and developers design robust and scalable systems. - Scale and optimize multiple databases - Build internal tooling that accelerates developer velocity - Provide observability, monitoring, and visibility across our systems ## On-Call Responsibilities (2–3 rotations per month) - You will participate in a shared on-call rotation, typically 2–3 times per month, covering the period from Friday at 7:00 AM ET through Saturday at 5:00 PM ET. ### How our on-call system works - You'll receive full internal training to understand our architecture, operational flows, and incident procedures - During your rotation, you are the primary point of escalation for production issues - On-call is not a daily responsibility, only during your assigned weekends - Shifts rotate across the team to maintain balance and avoid burnout ### What on-call requires - Strong understanding of system architecture and cross-service dependencies - Previous real-world experience in production on-call environments - Ability to quickly assess incidents, identify scope/root cause, and understand platform impact - Ability to classify severity and prioritize response appropriately - Capability to deploy safe production hotfixes when needed - Solid judgment under pressure - especially when operating independently - Ownership mindset: from detection to mitigation to resolution ## Requirements - 3+ years of experience as an independent backend or infrastructure engineer - Ability to design and build scalable, reliable systems - Strong communication skills - Hands-on builder mentality — this is a coding role. - Experience with: - Relational and non-relational databases - Major Cloud platforms (GCP, AWS, Azure), GCP is an advantage - Streaming systems - Scaling large systems - Message queues - Monitoring systems like DataDog, Grafana, Groundcover - CI/CD, Git ## Nice to Have - Bonus Skills: Kubernetes and Knative (production experience) & ClickHouse ## Technologies we use - Databases: ClickHouse, MongoDB, Redis, PubSub, Postgres, BigTable, Firestore, SonicDataFlow - Infrastructure: Kubernetes, Knative, GCP, Temporal, Pulumi - Languages: Node.js, TypeScript, Python - Tooling: Git, GitHub, GCP build infrastructure, GitHub Actions
Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Wellfound ↗