Platform Engineer - (Site Reliability Engineering)
Core
Own the full incident lifecycle for a high-scale crypto platform, including active response, postmortem automation, and root cause elimination.
Role type
Senior IC Platform Engineer (Site Reliability Engineering)
Builds
Incident management automation, observability dashboards, and internal tooling for a cryptocurrency exchange
Domain
Fintech / Cryptocurrency
Deliverable
production ML models | infrastructure
Required skills
Kubernetes, CI/CD pipelines, software development (Python/Java), incident management, automation, observability
Preferred skills
AI agents, LLM-based workflows, fintech/crypto domain knowledge
Technologies
Kubernetes, Python, Java
Responsibilities
Execute on-call shifts and drive incident resolution, build automation for postmortem workflows, leverage AI to identify incident patterns and propose systemic fixes, improve the observability ecosystem, participate in change management processes
Seniority
Senior, hands-on IC