Staff Engineer, Data Engineering
Core
Architect, design, and maintain scalable big data pipelines, ETL processes, and real-time analytics frameworks to process battery data and support AI/ML model deployment.
Role type
Staff Engineer, Data Engineering
Builds
Scalable cloud-based data infrastructure, real-time analytics frameworks, and production ML workflows for battery data processing.
Domain
Semiconductor / Battery Data / Cloud Data Engineering
Deliverable
production ML models | product features
Required skills
Big data pipeline architecture, ETL development, real-time analytics framework design, data quality validation, cloud data infrastructure, Python, Spark, analytics engineering, data warehousing, dimensional modeling, schema evolution, relational and non-relational database design, AWS services (S3, EC2, Glue, Lambda, Kinesis), Apache Kafka, Apache Spark, workflow orchestration (Airflow, Prefect), SQL, NoSQL, version control (GitHub), CI/CD automation, containerization (Docker), orchestration (Kubernetes), AWS SageMaker.
Preferred skills
Experience with time-series data from laboratory or device-based systems, open table formats (Apache Iceberg, Delta Lake), cross-functional collaboration with algorithm and hardware engineers.
Responsibilities
Architect and develop scalable big data pipelines and ETL processes; translate business objectives into functional pipeline specifications; identify and implement data storage and retrieval solutions; ensure data quality through validation and cleansing; deploy and support ML workflows in production; evaluate and recommend improvements to data infrastructure.
Seniority
Staff, hands-on IC with strategic coordination