Data Engineer Lead
Core
Design, build, and operate production-grade data pipelines and platforms using PySpark, Python, and SQL on AWS, integrating GenAI tooling to enhance development efficiency.
Role type
Senior hands-on IC Data Engineering Lead
Builds
Production data pipelines, data products, and scalable data platforms
Domain
Insurance/Financial Services, Cloud Data Engineering
Deliverable
production ML models | product features | infrastructure
Required skills
PySpark, Python, SQL, AWS EMR, AWS Glue, AWS S3, Snowflake, Data Modeling, Architecture Design, Code Review, Technical Documentation, GenAI Tooling
Preferred skills
AWS Solutions Architect certification, Insurance domain knowledge, ADR writing experience
Technologies
PySpark, Python, SQL, AWS EMR, AWS Glue, AWS S3, Snowflake, Claude Code, Snowflake Cortex
Responsibilities
Design and build production-grade PySpark and Python pipelines on AWS; Own end-to-end delivery of data products within Agile/Scrum; Translate business requirements into technical designs and data models; Set and enforce engineering standards and coding conventions; Conduct code reviews and mentor junior engineers; Leverage GenAI tooling to improve code quality and productivity; Produce technical documentation including ADRs and runbooks.
Seniority
Senior, hands-on IC with team leadership