CareerPlanSign in

Principal Engineer

Bengaluru, India💼 Full-time🗓 2026-07-10 → 2026-09-25

Core

Design, implement, and optimize production-grade Retrieval-Augmented Generation (RAG) pipelines to improve LLM accuracy and deploy GenAI applications on secure cloud infrastructure.

Role type

Senior IC Cloud Generative AI and RAG Engineer

Builds

Robust RAG pipelines, GenAI applications, and scalable cloud infrastructure

Domain

Generative AI, RAG, Cloud Infrastructure

Deliverable

production ML models

Required skills

Python (asynchronous programming, API development), GenAI frameworks (LangChain, LlamaIndex, Hugging Face), Vector Databases (Vespa, Pinecone, Milvus, Chroma), Cloud platforms (AWS, GCP, Azure), Docker, Kubernetes, Prompt Engineering

Preferred skills

Fine-tuning open-source LLMs (Llama, Mistral), MLOps/LLMOps tools (LangSmith, Weights & Biases), Cloud certifications

Technologies

FastAPI, LangChain, LlamaIndex, Hugging Face, Vespa, Pinecone, Milvus, Chroma, AWS, GCP, Azure, Docker, Kubernetes

Responsibilities

Design and optimize RAG pipelines for LLM accuracy; Write clean, modular backend code using Python; Deploy and scale GenAI applications on cloud platforms; Manage and tune vector databases for semantic search; Connect frontends and data systems with enterprise LLMs via APIs; Track LLM token usage, latency, and retrieval accuracy in production

Seniority

Senior, hands-on IC

Sourced via workday · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.