Data Engineer
Core
Design, develop, and optimize scalable ETL/ELT pipelines and cloud-native data processing solutions.
Role type
Data Engineer
Builds
Scalable data pipelines and cloud-based data platforms
Domain
Cloud data engineering
Deliverable
production ML models | infrastructure
Required skills
Python, PySpark, SQL, AWS services (S3, Glue, EMR, Lambda, Redshift, Athena), ETL/ELT frameworks, data warehousing concepts, data modeling, performance tuning
Preferred skills
Data orchestration tools (Airflow), CI/CD, DevOps practices
Technologies
AWS, PySpark, Airflow
Responsibilities
Design and maintain scalable ETL/ELT pipelines; Build and optimize data processing workflows using PySpark; Develop robust data solutions using Python and SQL; Integrate data from multiple sources and ensure data quality; Optimize data storage, processing, and query performance; Troubleshoot and resolve data pipeline and performance issues; Implement data governance, security, and best practices
Seniority
Mid-Senior, hands-on IC