Site Reliability Engineer
Core
Drive product and engineering department forward by ensuring reliability across the Paddle platform, acting as a glue between product teams to enable efficient software development and operations.
Role type
Senior IC Site Reliability Engineer (Platform)
Builds
Production-grade payment infrastructure, automation tooling, disaster recovery processes, and observability standards for a global SaaS platform.
Domain
Fintech / Payments / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Go, PHP Laravel, AWS ecosystem (ECS Fargate, SQS, EventBridge, RDS/Aurora), Terraform, Docker, Linux administration, microservices, distributed systems, monitoring (OpenTelemetry, Honeycomb, Grafana), CI/CD, debugging, performance tuning, cost optimization, GitOps.
Preferred skills
AI tools for development, security mindset, automation expertise.
Responsibilities
Develop and maintain tools to maximize engineering efficiency (deployment infrastructure, database upgrades); create and maintain disaster recovery processes; handle production incidents and author postmortems; run performance investigations and drive tuning; own cost optimization workstreams; advocate for GitOps methodology.
Seniority
Senior, hands-on IC