Generative AI Engineer
Core
Design, develop, and deploy AI-powered applications using Large Language Models (LLMs), RAG pipelines, and vector databases.
Role type
Generative AI Engineer (IC)
Builds
Production-grade AI applications and services
Domain
Generative AI, Large Language Models, Enterprise Software
Deliverable
production ML models
Required skills
Python, LLMs (OpenAI, Anthropic, Meta Llama, Google Gemini), RAG systems, Prompt engineering, Vector databases (Pinecone, Weaviate, Chroma, Milvus, FAISS), AI orchestration (LangChain, LlamaIndex), RESTful APIs, Cloud platforms (AWS, Azure, GCP)
Preferred skills
Fine-tuning open-source LLMs, AI agents, MLOps, Docker, Kubernetes, NLP, embeddings, CI/CD pipelines
Technologies
LangChain, LlamaIndex, Semantic Kernel, Haystack, Hugging Face Transformers, GitHub Actions, Jenkins, PostgreSQL, MongoDB, Redis
Responsibilities
Design and deploy LLM-powered applications; Build and optimize RAG pipelines; Develop prompt engineering strategies; Integrate vector databases; Fine-tune and evaluate foundation models; Build AI APIs; Implement AI orchestration workflows; Monitor AI application performance; Ensure responsible AI practices and compliance; Document AI architectures and workflows.
Seniority
Mid-Senior, hands-on IC