Data Engineer
Core
Design, build, and maintain high-volume data ingestion and processing pipelines in Databricks to transform external data into clean, structured inputs for downstream analysis and intelligence.
Role type
Mid-level Data Engineer
Builds
Production-grade data pipelines and Delta Lake storage structures
Domain
Cloud SaaS for energy, utility, telecom, and infrastructure
Deliverable
production ML models
Required skills
Databricks, Spark/PySpark, SQL, Delta Live Tables, Delta Lake, medallion architecture, schema evolution, data validation, pipeline orchestration, CI/CD
Preferred skills
French (spoken and written), Quebec residency
Technologies
Databricks, Delta Lake, Spark, PySpark, SQL
Responsibilities
Design and maintain ingestion pipelines from high-volume external APIs; Implement transformation workflows using medallion architecture; Configure and manage Delta Lake storage structures and optimization routines; Ensure pipeline reliability through monitoring, alerting, and error handling; Build and monitor workflows using Databricks Workflows or Delta Live Tables; Collaborate with Data Scientists to expose clean data for LLM/NLP pipelines.
Seniority
Mid-level, hands-on IC