Site Reliability Engineer - Data Engineering
Core
Operate and evolve large-scale data collection, storage, access, and query platforms to support trader analysis, simulation, reporting, and insights.
Role type
Site Reliability Engineer - Data Engineering
Builds
Large-scale data platforms (Kafka, HDFS, Dremio, in-house pipelines) serving traders and analysts
Domain
Financial services / Data Engineering
Deliverable
production ML models | infrastructure
Required skills
Linux system administration, distributed systems management, CI/CD automation, container orchestration (Kubernetes, Docker, Helm), data access technologies (Dremio, Presto), workflow orchestration (Airflow, Prefect), cloud platforms (AWS, GCP, Azure)
Preferred skills
Experience with multi-petabyte data infrastructure, knowledge of Kafka, HDFS, Kubernetes
Technologies
Kafka, HDFS, Dremio, Presto, Airflow, Prefect, Kubernetes, Docker, Helm, AWS, GCP, Azure
Responsibilities
Manage and monitor distributed systems and data processing platforms; drive systems automation and CI/CD; collaborate with engineers and traders to troubleshoot queries; evaluate and implement innovative technology solutions
Seniority
Mid-level, hands-on IC