Data Engineer
Core
Design, build, and operate scalable data pipelines and data flows enabling NBA decisioning across on-premise and cloud platforms.
Role type
Senior Data Engineer (Cloud Migration & Real-time Pipelines)
Builds
Scalable batch and near real-time data flows for MVP and future scale decisioning systems.
Domain
Sports Analytics / Big Data Engineering
Deliverable
production ML models
Required skills
Google Cloud Platform (BigQuery), Hadoop/Cloudera ecosystem, Apache Spark, Scala, Python, SQL, Data Modeling, CI/CD, Git, Data Integration (Pega CDH/Infinity), Distributed Systems, Data Quality & Governance, AI-assisted development tools.
Preferred skills
Experience with GitHub Copilot or similar AI coding tools, experience with Pega CDH/Infinity.
Technologies
Google Cloud Platform, BigQuery, Hadoop, Cloudera, Pega CDH, Pega Infinity, Apache Spark, Scala, Python, SQL, Git.
Responsibilities
Design and implement end-to-end data pipelines across on-premise and GCP; Build and maintain batch and near real-time data flows; Drive migration and modernization of data from Hadoop/Cloudera to GCP; Ensure reliable, scalable, and performant data delivery for decisioning and analytics use cases; Implement and maintain data integration with Pega CDH / Infinity; Implement monitoring, validation, and reconciliation processes for data pipelines.
Seniority
Senior, hands-on IC