Senior Site Reliability Engineer
Core
Own day-to-day administration of AWS accounts, databases, and a fully serverless platform, ensuring production health, security, and smooth operations.
Role type
Senior Site Reliability Engineer (Cloud/Serverless)
Builds
Production environment for a fully serverless platform
Domain
Cloud infrastructure (AWS) and Data Management
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
AWS serverless services, PostgreSQL administration, Infrastructure as Code (SST/Pulumi), Linux, Scripting (TypeScript/Python/bash), Production monitoring, Incident response
Preferred skills
Disaster recovery planning, Cost optimization, Team knowledge sharing
Technologies
AWS, Lambda, SQS, EventBridge, CloudWatch, S3, PostgreSQL, SST, Pulumi, Terraform, GitHub Actions, Docker
Responsibilities
Administer AWS services and accounts; manage database backups and disaster recovery; monitor production and address operational issues; lead debugging and incident response; refine infrastructure for deployability and scalability; share operational knowledge with the team
Seniority
Senior, hands-on IC