Site Reliability Engineer
Core
Design and build tools to automate secure machine provisioning, monitoring, metrics collection, and life cycle maintenance for a mission-critical High-Frequency Trading environment.
Role type
Site Reliability Engineer (SRE)
Builds
Automated provisioning tools, secrets management systems, and infrastructure reliability services
Domain
High-Frequency Trading (HFT), Distributed Systems, Low-Latency Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Python, Golang, Linux/UNIX systems administration, distributed systems debugging, automation, configuration management, CI/CD, networking protocols
Preferred skills
Security knowledge, fast learning ability, strong computer science fundamentals
Technologies
Python, Golang, Linux, Debian, CI/CD frameworks, Configuration management tools
Responsibilities
Design and build tools to automate secure machine provisioning, monitoring, metrics collection, and life cycle maintenance; Develop and maintain systems for secrets management, network configuration, and infrastructure reliability; Troubleshoot complex issues across Linux systems, including application, networking, OS, and kernel-level problems; Write high-quality software in Python or Golang to support systems engineering workflows; Build, deploy, and maintain new services using configuration management and automation frameworks
Seniority
Mid-to-Senior, hands-on IC