CareerPlanSign in

Senior Site Reliability Engineer

New York🌐 Remote💼 Full-time💰 $139,000–$139,000🗓 2026-08-07 → 2026-09-26

Core

Design, build, and operate large-scale, distributed, fault-tolerant systems and ML inference infrastructure for Adobe Stock and AI-powered products.

Role type

Senior Site Reliability Engineer (ML & Cloud Infrastructure)

Builds

Cloud-native, containerized systems, ML inference pipelines, and automation tooling for Adobe's creative community and AI services.

Domain

Cloud Infrastructure, Machine Learning Operations, Creative Technology

Deliverable

production ML models | infrastructure

Required skills

Python, Kubernetes, AWS (EC2 Auto Scaling Groups), vulnerability/patch management, distributed systems debugging, IaC (Terraform, Chef, Ansible), CI/CD (Jenkins, Argo CD), ML inference pipeline operations

Preferred skills

GPU-backed compute tuning, agentic AI workflows (LangGraph), relational database operations, Azure/GCP experience

Technologies

AWS, Kubernetes, SageMaker, Bedrock, vLLM, LangGraph, Aurora PostgreSQL, Memcached, Fastly, Datadome, New Relic, Splunk, Grafana, Prometheus, Terraform, Jenkins, Argo CD

Responsibilities

Design and operate large-scale distributed systems; manage patch and golden-image lifecycle; build and operate ML inference infrastructure; set infrastructure standards for agentic AI; partner across software and ML teams; handle on-call rotation for web services, databases, and ML systems.

Seniority

Senior, hands-on IC

Sourced via workday · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.