Data Engineer II, DASH Device Operations
Core
Design, build, and operate large-scale data infrastructure and pipelines to power analytics, AI/ML workloads, and intelligent automation for Device Operations.
Role type
Senior Data Engineer (AI/ML Infrastructure)
Builds
Scalable batch and real-time data pipelines, AI-ready datasets, and data warehouse schemas.
Domain
Device Operations & Supply Chain, Cloud Data Engineering, AI/ML Data Foundations
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Data modeling, ETL/ELT design, SQL optimization, Python programming, Distributed computing (Spark/Hive), AWS services (Redshift, S3, Kinesis, EMR, Lambda), Kafka, Data warehousing architecture
Preferred skills
Infrastructure as code, Ops automation (Chef/Puppet/Ansible), Machine learning system deployment, Technical mentoring
Technologies
AWS (Redshift, S3, Sagemaker, EMR, Kinesis, Lambda, EC2), Spark, Kafka, Hive, Hbase, Yarn, Python, SQL, Chef, Puppet, Ansible
Responsibilities
Design and implement scalable data pipelines for batch and real-time processing; Build and maintain data infrastructure supporting AI-ready datasets; Interface with technology teams to extract, transform, and load data from diverse sources; Implement data models and ETL/ELT processes using dimensional modeling or data vault approaches; Partner with scientists and engineers to ensure data infrastructure meets ML training and inference needs; Mentor junior data engineers on best practices and code quality.
Seniority
Senior, hands-on IC with mentorship responsibilities