Data Engineer, Forward Deployed Engineer
Core
Design, deploy, and refine high-performance data infrastructure, pipelines, and context engines to fuel next-generation AI, machine learning, and autonomous workflows in live customer environments.
Role type
Senior Forward Deployed Data Engineer
Builds
Scalable ETL/ELT pipelines, distributed database models, vector databases, feature stores, and API endpoints for AI agents
Domain
Enterprise technology modernization, AI/ML infrastructure, cloud-native data platforms
Deliverable
production ML models | infrastructure
Required skills
Python, SQL optimization, ETL/ELT orchestration (Airflow, dbt, Kafka), distributed systems architecture, vector databases (Pinecone, Milvus, Chroma, Weaviate), data observability (Great Expectations, DataHub), cloud platforms (AWS, Azure, GCP)
Preferred skills
Containerization (Docker, Kubernetes), CI/CD pipelines, semantic modeling, regulated industry experience (Financial Services, Healthcare, Public Sector)
Technologies
Airflow, dbt, Kafka, Pinecone, Milvus, Chroma, Weaviate, Great Expectations, DataHub, AWS, Azure, GCP, Docker, Kubernetes, Git, GitHub
Responsibilities
Design and optimize scalable ETL/ELT pipelines for batch and real-time streaming; Architect distributed systems and database models in hybrid/cloud environments; Deploy and manage vector databases and semantic layers for AI search and RAG; Implement data quality frameworks and pipeline observability; Build secure API endpoints for autonomous systems; Document deployment insights to evolve core platforms
Seniority
Senior, hands-on IC