Senior Data Engineer (with AI/ML experience) India
Core
Designing, building, and maintaining end-to-end data pipelines, architecture, and design for a warehouse, LLM-driven applications, and AI-based BI.
Role type
Senior Data Engineer (AI/ML focus)
Builds
Centralized feature stores, model-training workflows, real-time inference services, and LLM-powered applications
Domain
Financial services / Communications / AI & Machine Learning
Deliverable
production ML models
Required skills
Apache Spark, Hadoop, Kafka, SQL, PLSQL, Python, Snowflake, Redshift, vector databases, RAG architectures, open-source LLM frameworks, AWS/Azure Machine Learning, Unix/Linux/Windows, version control systems
Preferred skills
DataStax AstraDB, LangChain, LlamaIndex, Hugging Face Transformers, LLaMA-4, MLOps tooling, CI/CD pipelines
Technologies
Snowflake, Redshift, Apache Spark, Hadoop, Kafka, AWS, Azure, DataStax AstraDB, LangChain, LlamaIndex, Hugging Face
Responsibilities
Design scalable data pipelines for ingestion, transformation, and delivery into feature stores and model-training workflows; Optimize ETL/ELT pipelines and data structures within cloud data warehouses; Build workflows for extracting and storing semantic representations of unstructured data; Architect analytics and dashboarding solutions for natural language query experiences; Manage prompt engineering, orchestration flows, and model fine-tuning routines; Oversee vector data stores and develop indexing methodologies for RAG workflows; Partner with stakeholders to translate language-model initiatives into scalable solutions; Create and maintain documentation for data processes and model deployment routines
Seniority
Senior, hands-on IC