Data Engineer
Core
Build and maintain data pipelines that ingest, transform, and serve data for media and sustainability analytics products.
Role type
Data Engineer
Builds
Self-serve data platforms, data pipelines, and analytics-ready datasets
Domain
Media and sustainability analytics
Deliverable
production ML models | product features
Required skills
Apache Spark (PySpark), SQL, cloud platforms (Azure), Databricks, Unity Catalog, data governance (RBAC/ABAC), API-based data ingestion, CI/CD, version control
Preferred skills
Apache Airflow, identity/access management (Okta, Entra ID), service mesh (Istio), containerized deployments (AKS/Kubernetes), C4 model architecture documentation
Technologies
Azure, Databricks, Unity Catalog, PySpark, SQL, Power BI, Tableau, Apache Airflow, Kubernetes, AKS, Istio
Responsibilities
Build and maintain data pipelines ingesting from third-party APIs and internal sources; Develop data transformation logic to standardize and enrich raw data; Build and maintain orchestration workflows for pipeline execution; Support data governance and access control models; Work with product managers to implement platform features; Support integration with visualization tools; Troubleshoot data quality and pipeline failures; Work with DevOps/security teams on infrastructure migrations; Participate in on-call rotations
Seniority
Mid-level, hands-on IC