Web Crawling & Automation Internship (m/w/d)
Core
Develop and maintain Python-based web crawlers, monitor data quality, and automate reporting for data engineering tasks.
Role type
Intern, data engineering (web crawling & automation)
Builds
Automated data pipelines and reports
Domain
Data engineering, web technologies
Deliverable
production ML models (via careerplan.io/jobs/2676248-web-crawling-automation-internship-mwd-at-cheil-germany-gmbh)
Required skills
Python, SQL, HTML, APIs, JSON, web technologies, data cleaning, ETL/ELT concepts
Preferred skills
Airflow, Docker, Git/Bitbucket, AWS, Tableau, Excel
Technologies
Python, Requests, BeautifulSoup, Selenium, Playwright, Scrapy, SQL, Excel, Tableau, Airflow, Docker, Git, Bitbucket, AWS
Responsibilities
Support development and maintenance of Python-based web crawlers; Contribute to monitoring crawlers, data quality and ensuring smooth report automation; Help optimize SQL queries and troubleshoot data issues; Assist in data cleaning, preparation and validation process; Collaborate with data engineers, data analysts and data scientists to deliver usable data; Document technical processes and data-quality limitations.
Seniority
Intern
