CareerPlanSign in

Site Reliability Engineer

USA💼 Full-time💰 $71,000–$111,000🗓 2026-07-14 → 2026-10-03

Core

Design, deploy, and run highly available, fault-tolerant systems on cloud platforms to ensure scalability, security, and efficiency for global payment operations.

Builds

Highly available, disaster-ready systems and automated operational tooling for payment network infrastructure

Domain

Financial services / Cloud infrastructure / Observability

Required skills

Linux/Unix administration, scripting (Python, Go, Bash), cloud platform management (AWS, Azure, GCP), CI/CD pipeline design, containerization and orchestration, incident response, capacity planning, performance tuning, ITSM principles, observability tooling

Preferred skills

Splunk, Dynatrace, PCF platform operations, Jenkins, Bitbucket, XLR

Technologies

AWS, Azure, Bash, BitBucket, CI/CD, Dynatrace, GCP, Jenkins, Linux, Python, Splunk, Unix

Responsibilities

Implement and maintain high-availability systems, develop automation scripts to reduce manual work, troubleshoot and resolve system issues, document operating procedures, lead triage and root cause analysis, conduct blameless post-mortems, support production readiness and operational design, manage production changes, contribute to compliance and risk mitigation

Seniority

Mid-level, hands-on IC