Data Engineer
Core
Design, implement, and maintain scalable data pipelines, ETL processes, and APIs to ingest, catalog, normalize data, and ensure high data quality and availability for government agencies and commercial partners.
Role type
Senior Data Engineer
Builds
Scalable data-driven products, data pipelines, and APIs for government and commercial stakeholders
Domain
Government technology, cybersecurity, cloud data infrastructure
Deliverable
production ML models | product features | infrastructure
Required skills
Python, PySpark, SQL, AWS ETL tools (MSK, Firehose, SNS, Airflow), data modeling, data governance, metadata management, open table formats (Iceberg, Parquet), big data technologies (Hadoop, Spark), cloud services (AWS, GCP)
Preferred skills
Data mesh principles, domain ownership, data product thinking, federated governance, RBAC/ABAC models
Technologies
AWS (S3, Redshift, Athena, MSK, Firehose, SNS, Airflow), GCP, Hadoop, Spark, MySQL, NoSQL, Apache Iceberg, Parquet, DataZone
Responsibilities
Design and maintain scalable data pipelines and ETL processes; Create and manage APIs for data access and integration; Lead data modeling efforts to support business goals; Configure and maintain central data catalogs and governance platforms; Implement best practices for data governance and regulatory compliance; Collaborate with cross-functional teams to build scalable data-driven products
Seniority
Mid-level (3+ years experience)