Senior Site Reliability Engineer
Core
Ensure performance and resilience of robust backend systems across all product teams, focusing on incident response, observability, and secure coding practices.
Role type
Senior Site Reliability Engineer (Backend Systems)
Builds
Scalable backend services, shared infrastructure, and automated testing pipelines for a wealth management AI platform.
Domain
Fintech / Wealth Management / Cloud Infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Java, Spring Boot, RESTful APIs, Microservices architecture, Distributed systems, SQL (Postgres, MySQL), Docker, Kubernetes, Helm, Incident Response design, Open Telemetry, SLOs, Automated testing
Preferred skills
AWS, Terraform, Debezium, Temporal, GitLab CI/CD, Kafka, Snowflake, Tabletop exercises, AI best practices
Technologies
Java, Spring Boot, Postgres, MySQL, Docker, Kubernetes, Helm, Open Telemetry, Kafka, Temporal, Snowflake, Terraform, GitLab, AWS
Responsibilities
Mentor teams on secure coding practices; Implement best practice observability with Open Telemetry and SLOs; Own the Incident Response process design and followups; Employ automated testing practices to ensure high-quality code.
Seniority
Senior, hands-on IC