Intern – Data Engineer (Hybrid: Onsite & Remote)
Core
Develop and implement data pipeline solutions to support reporting, analytics, and data needs within the food distribution industry.
Role type
Summer intern data engineer
Builds
Large scale datasets and efficient data pipelines
Domain
Food distribution industry + Data Engineering
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Data modeling, SQL, Python, ETL, Cloud environments, Unit testing, Technical documentation
Preferred skills
AWS (ERM, Lambda, IAM, S3), Linux, CI/CD tools, Semi-structured data handling, LLM concepts, Agentic Development, Snowflake, Airflow, Spark Data Frames, Pandas, API development, Kafka
Technologies
SQL, Oracle, MySQL, Python, AWS, Linux, Chef, Jenkins, Bamboo, Git/Bitbucket, Snowflake, Airflow, Spark, Pandas, Mulesoft, Kafka, Tableau
Responsibilities
Design and build large scale datasets, Build, test and implement highly efficient data pipelines, Analyze new data sources to understand quality and content, Write technical design documentation, Create and execute unit tests, Support QA team during testing