Data Engineer (Familiar with LLM, ML or GenAI apps)
Core
Design, build, and maintain ETL/ELT pipelines and data models to power business insights and operationalize Generative AI applications.
Role type
Senior Data Engineer (GenAI/LLM focus)
Builds
Scalable data workflows, optimized data models, and GenAI inference pipelines
Domain
Government technology, cloud data infrastructure, Generative AI
Deliverable
production ML models
Required skills
Python, PySpark, SQL, AWS services (S3, Glue, Lambda, Redshift, Athena), Databricks, Tableau, Git, API design, data orchestration (Airflow, dbt), CI/CD
Preferred skills
ML/LLM application deployment, embedding stores, text/chat pipeline management
Technologies
Databricks, PySpark, AWS, Tableau, Airflow, dbt, Git
Responsibilities
Design and maintain ETL/ELT pipelines on AWS-based platforms; Develop and optimize data models for BI tools; Work with AI engineers to operationalize GenAI models including orchestrating data dependencies and inference pipelines; Implement monitoring, version control, and quality standards across data workflows; Collaborate with business teams to translate requirements into engineering solutions; Ensure data security and compliance with organizational guidelines
Seniority
Mid-to-Senior, hands-on IC