Senior Staff Software Development Engineer (Data Engineer)- Video Insights
Core
Architect scalable data pipelines and hybrid retrieval architectures (vector search + graph traversals) to power semantic search and RAG systems for AI agents.
Role type
Senior Staff Software Development Engineer (Data Engineer)
Builds
Production retrieval systems, knowledge graphs, and data infrastructure for GenAI platforms.
Domain
Artificial Intelligence / Semantic Search / Knowledge Graphs
Deliverable
production ML models
Required skills
Python, Java, Go, Vector Databases (HNSW, IVFFlat, PQ), Graph Databases (Cypher, Gremlin), LLM orchestration (LangChain, LlamaIndex), Spark, Flink, Kafka, Cloud-native infrastructure (AWS/GCP/Azure), Kubernetes
Preferred skills
MS or PhD in CS/Mathematics, 9+ years in Infrastructure/DevOps/SRE
Technologies
Pinecone, Milvus, Weaviate, Qdrant, Neo4j, AWS Neptune, ArangoDB, OpenAI, HuggingFace, Cohere, LangChain, LlamaIndex, Spark, Flink, Kafka, Kubernetes, HNSW, IVFFlat, PQ
Responsibilities
Lead design of hybrid retrieval architectures combining vector similarity search with structured graph traversals; Architect scalable data pipelines for ingestion, embedding, and indexing of massive multi-modal datasets; Innovate and prototype advanced retrieval techniques including multi-stage re-ranking and graph-tooling for LLMs; Design and implement schemas for complex knowledge graphs ensuring ontological integrity; Build automated data validation and drift detection systems; Drive technical implementation of "Memory" systems for AI agents; Collaborate with AI Research and Product teams to integrate emerging database technologies into production.