Site Reliability Engineer - Data Platform
Core
Design, implement, and operate distributed data platforms to support traders, quant researchers, and engineering teams.
Role type
Senior Site Reliability Engineer (Data Platform)
Builds
Foundational data platforms, observability systems, and automation for data frameworks and tooling.
Domain
Financial technology / Distributed data systems
Deliverable
production ML models | infrastructure
Required skills
Distributed data platforms management, Linux and Kubernetes orchestration, Infrastructure as Code, Python programming, SQL query tuning, JVM debugging, Data lakehouse technologies, Workflow orchestration
Preferred skills
None explicitly stated
Technologies
Kafka, HDFS, Spark, Dremio, Iceberg, Clickhouse, Airflow, Flink, Ansible, Puppet, Kubernetes, Prometheus, Grafana, AlertManager
Responsibilities
Design and operate data platforms, improve observability, build automation to reduce toil, support reliability of critical services, drive long-term architectural improvements
Seniority
Senior, hands-on IC