Intern, Data Engineer
Core
Build and optimize scalable data pipelines, ETL/ELT processes, and data models using Databricks and cloud platforms to enable data-driven decision-making.
Role type
Intern, Data Engineer
Builds
Modern data pipelines, analytics solutions, and scalable data workflows
Domain
Technology, Digital, and Analytics innovation hub for the insurance/healthcare sector
Deliverable
production ML models | product features
Required skills
Python, SQL, software development concepts, data engineering fundamentals, Git/GitHub, data pipeline design, ETL/ELT processes
Preferred skills
Databricks, Apache Spark, cloud platforms, AI-assisted development tools, Generative AI use cases, prompt engineering
Technologies
Databricks, Python, SQL, GitHub Copilot, Apache Spark
Responsibilities
Design, develop, and maintain data pipelines; Build and optimize ETL/ELT processes; Support development of scalable data models; Perform data quality validation and monitoring; Document data pipelines and technical solutions; Integrate data from databases, APIs, and files; Participate in testing and deployment activities; Contribute to proof-of-concepts involving Generative AI; Research emerging technologies and recommend innovative approaches.
Seniority
Intern