Data Engineer II, DSP Analytics
Core
Design and build next-generation data infrastructure and real-time/batch pipelines to power Generative AI applications and automate data operations for Amazon's Delivery Service Partner (DSP) program.
Role type
Senior Data Engineer (GenAI & Analytics Infrastructure)
Builds
Scalable data platforms, ETL pipelines, and automated data agents for GenAI workloads.
Domain
Logistics / Generative AI / Cloud Data Engineering
Deliverable
production ML models | infrastructure
Required skills
Data modeling, ETL pipeline development, AWS cloud technologies, modern programming languages (Python, Go, Java, etc.), SQL/NoSQL data analysis, system performance optimization, data governance.
Preferred skills
Non-relational databases (object, document, graph, column-family), AWS services (Redshift, S3, Glue, EMR, Kinesis, Lambda, IAM).
Technologies
AWS (Redshift, S3, Glue, EMR, Kinesis, FireHose, Lambda, IAM), Python, Ruby, Golang, Java, C++, C#, Rust, NoSQL, Oracle.
Responsibilities
Lead and design complex data architectures ensuring scalability and security; Drive implementation of large-scale data pipelines and ETL processes; Identify and resolve system-wide data challenges and performance bottlenecks; Establish data engineering best practices for governance and operational excellence; Ensure data solutions are auditable and maintainable; Collaborate with cross-functional teams to deliver innovative data solutions.
Seniority
Mid-Senior, hands-on IC