CareerPlanSign in
💼 Full-time🗓 2026-06-25

Core

Own the end-to-end LLM pipeline from retrieval architecture to fine-tuning and production deployment for a 0-1 AI product.

Role type

Senior IC machine-learning engineer (LLM/RAG)

Builds

Retrieval-Augmented Generation systems and LLM pipelines

Domain

Artificial Intelligence / Large Language Models

Deliverable

production ML models

Required skills

LLMs (Llama, Mistral, GPT-4, Claude), LangChain, LlamaIndex, hybrid search, reranking, vector databases (Pinecone, Weaviate), LoRA/QLoRA, Hugging Face Transformers, Python, FastAPI, PyTorch, Redis, Postgres, Docker, Kubernetes, automated evaluation loops

Preferred skills

IoT/Hardware integration, real-time systems, high-performance startups, defense-related AI projects

Technologies

Llama, Mistral, GPT-4, Claude, LangChain, LlamaIndex, Pinecone, Weaviate, FastAPI, PyTorch, Redis, Postgres, Docker, Kubernetes

Responsibilities

Build and maintain the full LLM lifecycle (Retrieval, Evaluation, Fine-tuning); Design and optimize RAG systems using hybrid search; Make critical architecture decisions for scale and sub-second speed; Monitor and improve accuracy and evaluation metrics; Translate business needs into technical AI roadmaps

Seniority

Senior, hands-on IC

Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.