Data Engineer Scala
Core
Design and maintain scalable, resilient data pipelines in distributed environments, ensuring data governance and quality to support technical architecture decisions.
Role type
Data Engineer
Builds
Scalable data pipelines and data infrastructure
Domain
Cloud data engineering (AWS)
Deliverable
production ML models | product features | infrastructure
Required skills
Apache Spark/Pyspark, Scala, AWS (Glue, S3, EMR, Athena, Redshift), data modeling, ETL/ELT, batch and streaming processing, Apache Kafka or Amazon MSK
Preferred skills
Legacy system modernization, Delta Lake, Apache Hudi, Iceberg, CI/CD for data (dbt, Airflow, Terraform)
Technologies
Apache Spark, Scala, AWS, Apache Kafka, Amazon MSK, Delta Lake, Apache Hudi, Iceberg, dbt, Airflow, Terraform
Responsibilities
Design, build, and maintain scalable and resilient data pipelines; Work with large volumes of structured and unstructured data; Ensure data quality, consistency, and governance; Collaborate with software engineers, analysts, and data scientists; Participate in technical decisions regarding data architecture and tools
Seniority
Mid-level to Senior, hands-on IC