Data Engineer
Core
Design and maintain scalable ETL pipelines and data engineering systems to support machine learning and data science teams.
Role type
Data Engineer (ETL & Infrastructure)
Builds
ETL pipelines, data processing systems, and engineering infrastructure for ML workloads.
Domain
Data Engineering, Machine Learning Infrastructure, Time-series Analytics
Deliverable
production ML models | infrastructure
Required skills
Large-scale data processing (Spark, Hadoop), Python/Scala/Go, Spark cluster management and tuning, DAG orchestration (Airflow), Open-source contribution
Preferred skills
Terraform, Docker, Kubernetes, Data engineering infrastructure management
Responsibilities
Design scalable ETL systems, identify and resolve scalability bottlenecks, manage Spark clusters, contribute to engineering infrastructure, improve data quality and discoverability
Seniority
Mid-level, hands-on IC