Data Engineer
Core
Designing, building, and tuning high-performance data integration pipelines and services for a bank, handling large data volumes (~100TB) on-premises and in the cloud.
Role type
Senior Data Engineer (ETL & Cloud)
Builds
High-performance data integration pipelines, REST API services, and cloud data infrastructure
Domain
Banking, Data Engineering, Cloud (GCP), On-premises
Deliverable
production ML models | product features | infrastructure
Required skills
Python, PySpark, Apache Airflow, GCP Data Flow, GCP Data Proc, Oracle, PostgreSQL, ScyllaDB, Kafka, GCP Pub/Sub, Linux, REST API development, NIFI, Automate Now
Preferred skills
Java (MicroServices), Groovy, Apache JMeter, Grafana, GIT
Technologies
Informatic Power Center, NIFI, Oracle, PostgreSQL, ScyllaDB, Automate Now, Apache Airflow, GCP Data Flow, GCP Data Proc, GCP BigQuery, GCP BigTable, Scylla Cloud, Kafka, GCP Pub/Sub, PySpark, Rust, Linux, JAVA, GIT, Grafana, Apache JMeter
Responsibilities
Designing and tuning relational and NoSQL databases (Oracle, PostgreSQL, ScyllaDB) for performance; Building and maintaining ETL pipelines on-premises (Informatic Power Center, NIFI) and in GCP (Airflow, Data Flow, Data Proc); Developing and exposing REST API services; Managing data ingestion via Kafka and GCP Pub/Sub; Ensuring 24/7 high-performance data service availability; Writing efficient data loading tools using Rust and Python.
Seniority
Senior, hands-on IC
