CareerPlanGet AI match score →

Senior DevOps Engineer – AI-Native Healthcare Platform

Onsite or remote • Pune+1🌐 Remote💼 Full-time🗓 2026-06-25

Core

Design and run scalable, secure infrastructure for an AI-native, multi-tenant SaaS platform serving healthcare needs.

Role type

Senior DevOps Engineer (AI-Native Healthcare)

Builds

Scalable cloud infrastructure, CI/CD pipelines, containerized AI workloads, and observability systems.

Domain

Healthcare technology, Cloud Infrastructure, AI/ML Systems

Deliverable

production ML models | infrastructure

Required skills

AWS (or Azure/GCP), Docker, Kubernetes, CI/CD pipeline design, Terraform/CloudFormation, Python/Shell scripting, distributed systems architecture, incident response, observability (Prometheus, Grafana, ELK)

Preferred skills

Experience with AI/data workloads, low-latency system optimization

Technologies

AWS, GitHub Actions, GitLab, Jenkins, Docker, Kubernetes, Prometheus, Grafana, ELK, Terraform, CloudFormation, Python, Shell

Responsibilities

Design and run scalable infrastructure on AWS; Build and improve CI/CD pipelines; Work with Docker and Kubernetes for deployment and scaling; Implement monitoring, logging, and alerting; Support AI workloads with low latency and high concurrency; Use Infrastructure as Code to build reproducible infrastructure; Implement secure infrastructure practices and ensure healthcare-grade compliance.

Seniority

Senior, hands-on IC

Rewrite
## About the role You'll own the infrastructure that everything depends on Not just uptime—but: - Systems that scale unpredictably. - Deployments that are fast and safe. - Failures that are handled before they escalate. ## Responsibilities ### Cloud & Systems - Design and run scalable infrastructure on AWS (or similar). - Support multi-tenant, high-throughput SaaS environments. ### CI/CD & Automation - Build and improve CI/CD pipelines (GitHub Actions, GitLab, Jenkins). - Automate build, test, and deployment workflows. ### Containers & Orchestration - Work with Docker and Kubernetes for deployment and scaling. - Ensure system stability under load. ### Observability & Reliability - Implement monitoring, logging, and alerting (Prometheus, Grafana, ELK). - Diagnose and resolve production issues. ### AI-Native Infrastructure - Support AI workloads (low latency, high concurrency). - Enable reliable deployment of AI-powered services. ### Infrastructure as Code - Use Terraform / CloudFormation. - Build reproducible, version-controlled infrastructure. ### Security & Compliance - Implement secure infrastructure practices. - Ensure healthcare-grade reliability and compliance. ## Requirements - Strong hands-on experience with AWS (or Azure/GCP), Docker, Kubernetes - Experience running production SaaS systems at scale - CI/CD, observability, and incident response experience - Infrastructure as Code (Terraform, etc.) + scripting (Python/Shell) - Understanding of distributed systems and system reliability - Exposure to AI/data workloads is a plus. ## What we offer - Own infrastructure that directly impacts real-world care delivery - Work on AI-native systems at scale - High ownership, real problems, no busywork ## About the company - Predictable infrastructure work → not this - Ticket-based ops → not this
Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Wellfound ↗