Senior AI/ML Ops Engineer-II (Hybrid in Bangalore)
Core
Designing, developing, and overseeing scalable AI/ML Ops platforms and pipelines to deploy and manage AI agents, foundation models, and RAG stacks for enterprise SaaS solutions.
Role type
Senior AI/ML Ops Engineer
Builds
Scalable AI/ML Ops platforms, model deployment pipelines, and infrastructure for training and serving AI agents and foundation models.
Domain
Enterprise SaaS, AI/ML Operations, Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
System design, AI/ML frameworks (LangChain, LangGraph), MLOps workflows (Databricks, MLflow, Mosaic AI Agent Framework), Cloud platforms (AWS, Azure, GCP), Python, SQL, Kubernetes, CI/CD, Infrastructure as Code (Terraform), Observability, Vector Search, Knowledge Graphs, Data governance.
Preferred skills
Monte Carlo for monitoring, AWS Bedrock, Unity Catalog, Databricks Lakehouse ecosystem, AWS hosted data platforms.
Technologies
Databricks, MLflow, Mosaic AI Agent Framework, Unity Catalog, Vector Search, Knowledge Graph, LangChain, LangGraph, Kubernetes, Terraform, Monte Carlo, AWS Bedrock, AWS, Azure, GCP.
Responsibilities
Design and implement automated CI/CD pipelines for model deployment; Provision and optimize infrastructure for training and serving; Implement post-deployment monitoring for model performance and data drift; Automate retraining and data pipeline workflows; Manage deployment of foundation models and RAG stacks; Manage GPU/CPU utilization to minimize cloud costs; Collaborate with data scientists and engineers to bridge model development and production; Manage versioning for data, code, and models; Implement data security measures and compliance policies.
Seniority
Senior, hands-on IC