CareerPlanGet AI match score →

Senior AI Engineer

Bengaluru, Karnataka, India💼 Full-time🗓 2026-04-14 → 2026-07-31

Core

Design, train, optimize, and deploy domain-specialized language models (prompts and fine-tuned SLMs) for enterprise automation in healthcare and insurance.

Role type

Senior AI Engineer (LLM/Prompt Engineering & SLM)

Builds

Production-grade prompt architectures, fine-tuned Small Language Models (SLMs), and RAG pipelines for enterprise workflows.

Domain

Healthcare and Insurance / Enterprise Automation / Regulated AI Systems

Deliverable

production ML models

Required skills

Prompt Engineering for 7B–13B models, LLM fine-tuning (LoRA, QLoRA, adapters), model distillation, PyTorch engineering, transformer architecture understanding, GPU-based training/inference, Docker/Kubernetes deployment, RAG architectures, data privacy and AI auditability design

Preferred skills

Experience with 8B-class models (LLaMA, Mistral, Qwen), building evaluation datasets and benchmarking frameworks, instruction tuning and supervised fine-tuning (SFT) pipelines, multimodal system collaboration

Technologies

LLaMA, Mistral, Qwen, PyTorch, Hugging Face, Accelerate, DeepSpeed, Triton, Docker, Kubernetes

Responsibilities

Design production-grade prompt architectures for 8B-class models; Develop structured prompts for enterprise tasks such as classification, extraction, reasoning, and summarization; Optimize prompts for accuracy, latency, and cost efficiency; Build prompt evaluation frameworks to measure accuracy, hallucination rates, and consistency; Design reusable prompt libraries and prompt templates for enterprise workflows; Develop prompt-to-model migration strategies converting high-performing prompts into fine-tuned SLMs; Design and fine-tune LLMs for domain-specific enterprise tasks; Develop Small Language Models (SLMs) optimized for enterprise deployment; Build instruction tuning and supervised fine-tuning (SFT) pipelines; Design evaluation datasets and automated benchmarking frameworks; Implement retrieval augmented generation (RAG) pipelines and tool-augmented workflows; Collaborate with speech AI and document AI teams to build multimodal systems; Deploy models in private cloud or on-premise environments with strong security controls

Seniority

Senior, hands-on IC

Sourced via workable · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Workable ↗