Staff Site Reliability Engineer, Database
Core
Ensure reliability, scalability, and performance of multi-terabyte scale PostgreSQL clusters and systems for a fintech trading platform.
Role type
Staff Site Reliability Engineer (Database)
Builds
High-availability, high-performance PostgreSQL database infrastructure and low-latency trading systems
Domain
Fintech / Trading / Database Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
PostgreSQL (multi-terabyte scale), Go, Prometheus, Linux, distributed tracing, schema design, query optimization, capacity planning, incident management, SLI/SLO/SLA design
Preferred skills
pgx, gorm, sqlc, low-latency system experience
Technologies
PostgreSQL, Go, Prometheus, Linux, pgx, gorm, sqlc
Responsibilities
Triage difficult technical problems and implement solutions; Improve observability stack (monitoring, logging, profiling); Respond to and resolve incidents and conduct post-incident reviews; Collaborate with development teams on reliability and scalability; Monitor system capacity and implement changes for future growth
Seniority
Staff, hands-on IC