Generative AI Engineer
Core
Design and build bespoke Generative AI solutions for enterprise clients by orchestrating LLMs, RAG pipelines, and autonomous agents to solve specific business problems.
Role type
Senior Generative AI Engineer (Consulting)
Builds
Production-ready GenAI applications, RAG pipelines, multimodal agents, and prompt engineering frameworks for enterprise customers.
Domain
Management Consulting / Generative AI / Large Language Models
Deliverable
production ML models | product features
Required skills
Python, PyTorch, Hugging Face, LangChain, LlamaIndex, Prompt Engineering, RAG, Vector Databases (Pinecone, FAISS, pgvector), Cloud AI Services (AWS SageMaker, GCP Vertex AI, Azure Bedrock), MLOps (CI/CD, drift detection), Fine-tuning (PEFT, LoRA), API development (FastAPI/Flask), Transformer architectures.
Preferred skills
Experience with commercial vs open-source LLM benchmarking, Autonomous agent flows, Multimodal AI (text, audio, image), Responsible AI compliance frameworks.
Technologies
OpenAI, Claude, Mistral, LangChain, LlamaIndex, Pinecone, FAISS, pgvector, ChromaDB, Qdrant, Milvus, Weaviate, AWS SageMaker, GCP Vertex AI, Azure Bedrock, FastAPI, Flask, PyTorch, Hugging Face.
Responsibilities
Design and build robust GenAI applications using LLMs and frameworks like LangChain; Implement RAG pipelines with vector databases for grounding LLM responses; Develop multimodal AI solutions and autonomous agents; Drive MLOps excellence including CI/CD, drift detection, and retraining schedules; Design reusable prompt templates and optimize prompt flows; Deploy GenAI models on major cloud platforms; Ensure performance observability, security guardrails, and compliance; Benchmark LLMs for cost-performance tradeoffs; Evaluate fine-tuning strategies for proprietary use cases.
Seniority
Senior, hands-on IC