CareerPlanGet AI match score →

Team Lead Software Systems Engineering

Bengaluru, Karnataka, India💼 Full-time🗓 2026-06-28 → 2026-08-03

Core

Ensuring stability and reliability of the Reservations suite of products through SRE practices, automated delivery, and incident management.

Role type

Lead Site Reliability Engineer (SRE)

Builds

Production-ready infrastructure, CI/CD pipelines, and observability solutions for the Reservations suite

Domain

Travel technology, Cloud-native systems

Deliverable

infrastructure

Required skills

Google Cloud Platform, Linux/UNIX, Terraform, shell scripting, networking, CI/CD concepts, Jenkins, Docker, Kubernetes, monitoring and alerting tools (AppDynamics, Google Cloud ops, Data dog, Prometheus, Grafana, Elastic search), analysis and debugging, change management

Preferred skills

Java development, relational databases (Oracle, Couchbase, Datastore, Spanner), Google SRE knowledge, Service Now, runbooks

Technologies

GCP, Jenkins, Docker, Kubernetes, Terraform, AppDynamics, Data dog, Prometheus, Grafana, Elastic search, Oracle, Couchbase, Datastore, Spanner

Responsibilities

Provide application on-call support and troubleshoot system alerts, lead investigation of severe incidents, build alerting and monitoring solutions, continuously improve reliability via SRE practices, support development partners with infrastructure deployments, take ownership of services including capacity monitoring and cost optimization, provide technical mentorship

Seniority

Lead, hands-on IC with mentorship

Full job description

Powering the agentic revolution in travel. Sabre is an AI-native technology leader, backed by one of the world’s largest travel data clouds. Built on an open, modular, cloud-native architecture, Sabre serves as the backbone for both established leaders and bold, new disruptors, guiding them to the next age of travel retailing through intelligent, connected, and personalized experiences. With AI at its core and operating at unparalleled scale, Sabre transforms insights into innovation, empowering airlines, hoteliers, agencies and other partners to retail, distribute and fulfill travel worldwide.

Sabre is seeking a talented Lead Site Reliability Engineer to support the PNR SRE Team in Bangalore. This team supports the operations of the "Reservations" suite of products at Sabre and is globally distributed across US, Poland and India. We are looking for a team member to join this team and be focused on ensuring the stability and reliability of the services, automated delivery, increasing velocity of releases, and working with us on incorporating best practices, high quality and ensuring overall great customer experience for Sabre customers. The team member must practice SRE principles and should strive to reduce toil and treat reliability as a feature. This role will report to the team manager.

Role and Responsibilities:

• Provide application on-call support (8–12-hour shifts, one week at a time), responding to system alerts/pages, troubleshooting issues, and taking action when needed.

• Participate on incident bridges and lead investigation of severe issues affecting customer experience.

• Review, execute, validate and monitor changes to ensure stability and risk are addressed

• Build alerting and monitoring solutions for new or existing services in GCP, and other observability tools.

• Continuously improve reliability by engaging in SRE practices such as blameless postmortems, building SLIs for CUJs, reducing toil and increasing automation.  

• Support development partners with infrastructure deployments, ci-cd pipelines, and other non-functional requirements to be prod-ready

• Take ownership of services, addressing risks, and keeping up with other KTLO tasks including managing infrastructure currency, PCI audits, capacity monitoring, and cost optimization.

• Provides technical mentorship and cultural/competency-based guidance to teams.

Qualifications and Education Requirements:

• Google Cloud Platform knowledge

• Analysis, debugging and troubleshooting skills, persistence in problem solving

• Good collaboration within the team and with teams up and across the organization

• Understanding of CI/CD concepts

• Familiarity with Jenkins

• Understanding of container orchestration technologies (Docker, Kubernetes)

• Linux/UNIX knowledge, Terraform, shell scripting, networking

• Familiarity with monitoring and alerting tools: AppDynamics, Google Cloud ops, Data dog, Prometheus, Grafana, Elastic search 

• Experience with Change Management process

• Very good written and verbal English communication skills

• Willingness to learn 

• Self-disciplined and commitment oriented

NICE TO HAVE SKILLS:

• Understanding of databases (relational, Oracle, Couchbase, Datastore, Spanner)

• Java development experience

• Experience with Sabre Change Management process, Service Now, runbooks

• Google SRE knowledge

We will give careful consideration to your application and review your details against the position criteria. You will receive separate notification as your application progresses.

Please note that only candidates who meet the minimum criteria for the role will proceed in the selection process.

#LI-Hybrid#LI-NG1

Sourced via workday · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Workday ↗