Staff Site Reliability Engineer (founding) at Talonic
Core
Establishing the first dedicated Site Reliability Engineering function to ensure 99.9% availability, manage SLOs, and oversee compliance-driven reliability for enterprise document ingestion systems.
Role type
Staff Site Reliability Engineer (founding)
Builds
Enterprise document ingestion platform with canonical field registry and schema projection
Domain
Enterprise document management, cloud infrastructure, compliance (GDPR, ISO 27001/42001, HIPAA)
Deliverable
production ML models | infrastructure
Required skills
SLO definition and error-budget management, observability architecture, post-mortem analysis, deployment pipeline design, compliance audit integration, AWS infrastructure, Docker, Railway
Preferred skills
Claude Code integration, founding team experience, owning Sev1 incidents
Technologies
AWS, Docker, Railway, Claude Code
Responsibilities
Define and enforce SLOs and error-budget policies, build observability systems for critical paging, design post-mortem processes that drive architectural changes, manage deployment pipelines to reduce fear of releases, integrate reliability practices with compliance audits
Seniority
Staff, founding IC