Data Engineer - GenAI
Core
Build specialized data infrastructure and pipelines to power Generative AI, RAG systems, and autonomous agents for global clients.
Role type
Senior Data Engineer (GenAI)
Builds
GenAI-ready production pipelines, scalable AI platforms, vector databases, and data-as-a-service layers.
Domain
Generative AI, Data Engineering, Cloud Infrastructure
Deliverable
production ML models
Required skills
Python, PySpark, SparkSQL, dbt, Airflow, Docker, Kubernetes, SQL, Git
Preferred skills
LLM fine-tuning, RAG architecture, vector data management, LLMOps
Technologies
Databricks, Azure, AWS, dbt, Airflow, Prefect, Dagster, Docker, Kubernetes
Responsibilities
Design and deploy batch and streaming pipelines for unstructured data; Develop scalable AI platforms and vector DBs; Implement data governance and transformation best practices; Collaborate with ML engineers on data models for AI agents; Advocate for clean code and microservices deployment.
Seniority
Mid-Senior, hands-on IC