Site Reliability Engineer
Core
Architect, build, and maintain world-class production infrastructure for a global trading system, ensuring reliability, performance, and operability.
Role type
Senior Site Reliability Engineer (Production Infrastructure)
Builds
Production trading system, orchestration, configuration management, and monitoring automation
Domain
Financial technology / High-frequency trading infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Linux systems engineering, Python scripting, Shell scripting, Network engineering, System performance tuning, Incident management, Configuration management, DevOps tooling, Risk assessment
Preferred skills
C++ programming
Technologies
Linux, Python, Shell, Networking protocols (routing, multicast, LLDP, VLAN, ethernet)
Responsibilities
Proactively monitor and troubleshoot large-scale trading systems and exchange connectivity; Build and maintain DevOps toolkit for deployment and monitoring; Collaborate with traders and Risk Management to coordinate changes and manage incidents; Assess and manage operational risk of production changes; Provide mentorship and cross-training to other SREs
Seniority
Senior, hands-on IC