Data Platform Engineer
Core
Design and build enterprise knowledge foundations enabling accurate, governed, and context-aware AI and agentic systems, including production RAG engineering and knowledge graphs.
Role type
Senior IC data platform engineer (AI/LLM infrastructure)
Builds
Production RAG systems, knowledge graphs, ingestion/transformation/indexing/retrieval pipelines, and governed data platforms.
Domain
AI/LLM infrastructure, data engineering, knowledge management, vector search
Deliverable
production ML models | product features
Required skills
Data engineering, Python, SQL, distributed processing, orchestration, vector search, hybrid retrieval, knowledge graphs, metadata modeling, semantic modeling, document parsing, chunking, taxonomy design, ontology design, entity resolution, lineage tracking, access control, CI/CD, observability, performance engineering
Preferred skills
Advanced proficiency with Databricks Assistant, GitHub Copilot, Cursor, LlamaIndex, LangChain, Neo4j, Pinecone, Weaviate, Elasticsearch/OpenSearch, cloud vector-search services
Technologies
Databricks, GitHub Copilot, Cursor, LlamaIndex, LangChain, Neo4j, Pinecone, Weaviate, Elasticsearch, OpenSearch
Responsibilities
Architect ingestion, transformation, indexing, retrieval, and knowledge-enrichment pipelines; Build production RAG systems using structured, unstructured, graph, and metadata-driven retrieval; Define document parsing, chunking, taxonomy, ontology, entity resolution, and lineage strategies; Implement access-aware retrieval, freshness controls, quality checks, and auditability; Optimize retrieval quality, latency, scale, and cost; Partner with AI engineers on context construction and grounding; Mentor engineers and define reusable knowledge-engineering standards
Seniority
Senior, hands-on IC