Senior Data Engineer
Core
Design, develop, and deploy high-performance data pipelines (streaming, micro-batch, batch) to build aggregate and enriched data sets for business analysis and visualization.
Role type
Senior IC data engineer
Builds
Data pipelines, data lakes, data warehouses, and data sets
Domain
Cloud data engineering
Deliverable
production ML models | product features | dashboards & analysis
Required skills
Data pipeline architecture, SQL, Python, Google Cloud Platform (BigQuery, GCS, Pub/Sub, Data Flow), workflow orchestration (Airflow, Oozie), data format handling (JSON, Parquet, ORC, XML), query performance tuning
Preferred skills
Real-time data ingestion (Kafka, Spark, Storm), NoSQL databases, data visualization tools (Looker, Tableau, PowerBI)
Technologies
Google Cloud, BigQuery, GCS, Pub/Sub, Data Flow, Airflow, Oozie, Kafka, Spark, Storm, JSON, Parquet, ORC, XML
Responsibilities
Own and drive data integration architecture; Design, develop, and deploy high-performance data pipelines; Build aggregate and enriched data sets; Document, maintain, and support data pipelines; Measure and monitor data quality and availability; Implement data policies (retention, privacy, security, compliance); Mentor junior data engineers
Seniority
Senior, hands-on IC