Engineer - Site Reliability
Core
Ensuring high-availability, low-latency operations for real-time trading platforms across European and US markets through infrastructure management, incident response, and automation.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Real-time low-latency trading platforms and operational support systems
Domain
Financial markets infrastructure / Trading technology
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Linux OS tuning, kernel-bypass networking stacks (Solarflare/Onload), Python, SQL, UNIX Shell, capacity planning, incident triage, root cause analysis, configuration management, automation development
Preferred skills
Cloud platforms (AWS, Azure, GCP), containerization (Docker, Kubernetes), C++, familiarity with US equities/options/futures market structures
Technologies
Solarflare, Onload, Linux, Python, SQL, Docker, Kubernetes, AWS, Azure, GCP
Responsibilities
Monitor and maintain low-latency bare-metal infrastructure; respond to production incidents during US market hours; perform root cause analysis and post-incident reviews; execute daily change tickets for production updates; analyze technical data sets to troubleshoot issues; develop Python tools for system health automation; participate in weekend testing and capacity planning.
Seniority
Senior, hands-on IC