Principal Engineer
Core
Design, implement, and optimize production-grade Retrieval-Augmented Generation (RAG) pipelines to improve LLM accuracy and deploy GenAI applications on secure cloud infrastructure.
Role type
Senior IC Cloud Generative AI and RAG Engineer
Builds
Robust RAG pipelines, GenAI applications, and scalable cloud infrastructure
Domain
Generative AI, RAG, Cloud Infrastructure
Deliverable
production ML models
Required skills
Python (asynchronous programming, API development), GenAI frameworks (LangChain, LlamaIndex, Hugging Face), Vector Databases (Vespa, Pinecone, Milvus, Chroma), Cloud platforms (AWS, GCP, Azure), Docker, Kubernetes, Prompt Engineering
Preferred skills
Fine-tuning open-source LLMs (Llama, Mistral), MLOps/LLMOps tools (LangSmith, Weights & Biases), Cloud certifications
Technologies
FastAPI, LangChain, LlamaIndex, Hugging Face, Vespa, Pinecone, Milvus, Chroma, AWS, GCP, Azure, Docker, Kubernetes
Responsibilities
Design and optimize RAG pipelines for LLM accuracy; Write clean, modular backend code using Python; Deploy and scale GenAI applications on cloud platforms; Manage and tune vector databases for semantic search; Connect frontends and data systems with enterprise LLMs via APIs; Track LLM token usage, latency, and retrieval accuracy in production
Seniority
Senior, hands-on IC