Senior Data Engineer (AI/ML)
Core
Design and build scalable data platforms and pipelines powering Generative AI products, including RAG systems, LLM applications, and vector search.
Role type
Senior IC data engineer (Generative AI/LLM infrastructure)
Builds
Production-grade AI/LLM data pipelines, RAG systems, semantic search, and data platforms for AI applications
Domain
Hospitality technology + Generative AI / Large Language Models
Deliverable
production ML models
Required skills
Python, Scala/Java, advanced SQL, Databricks, Snowflake, Apache Spark, Delta Lake, Airflow, LLM application development, RAG architecture, embeddings, vector databases, semantic search
Preferred skills
Kafka, Spark Structured Streaming, real-time data architectures, LangGraph, LangChain, LlamaIndex, Qdrant, Pinecone, Weaviate, Databricks Vector Search, MLflow, Unity Catalog, AI observability frameworks
Technologies
Databricks, Snowflake, Apache Spark, Delta Lake, Airflow, Kafka, LangGraph, LangChain, LlamaIndex, Qdrant, Pinecone, Weaviate, MLflow, Unity Catalog
Responsibilities
Design AI/LLM data pipelines for training, inference, and evaluation; Build production RAG systems with ingestion, chunking, and retrieval; Develop AI applications using LLMs and agentic workflows; Optimize semantic search and vector retrieval systems; Establish data quality, governance, and observability practices; Partner with ML teams to move prototypes to production
Seniority
Senior, hands-on IC