Staff Data Engineer
Core
Design and scale data infrastructure and pipelines powering analytics, machine learning, and business intelligence for a software supply chain security company.
Role type
Staff Data Engineer (Infrastructure & Platform)
Builds
Scalable data pipelines, data lakehouse architecture, and trusted datasets for analytics and ML.
Domain
Software Supply Chain Security / Data Engineering
Deliverable
production ML models | dashboards & analysis | infrastructure
Required skills
Distributed data systems (Spark, Kafka), ETL/ELT pipeline design, Data modeling (star schema, dimensional), SQL/NoSQL optimization, Python/Scala/Java, Cloud data platforms (AWS), Data observability, CI/CD for data
Preferred skills
Databricks optimization (Delta Lake), AI-assisted development tools, Workflow orchestration (Airflow, Dagster), Real-time/streaming architectures, Modern table formats (Iceberg, Hudi), Software supply chain domain knowledge
Technologies
Databricks, Spark, Kafka, AWS, Delta Lake, Airflow, Dagster, Python, Scala, Java, SQL, NoSQL
Responsibilities
Design, build, and maintain scalable data pipelines and ETL/ELT processes; Architect and optimize data models and storage solutions; Collaborate with data scientists and engineers to deliver trusted datasets; Own and evolve data platform components; Implement observability, alerting, and data quality monitoring; Drive best practices in data engineering; Mentor team members on engineering best practices; Contribute to next-generation data lakehouse architecture design.
Seniority
Staff, hands-on IC with architectural vision and mentorship