CareerPlanGet AI match score →

Senior Site Reliability Engineer (SRE) – DevOps & Production Support

Onsite or remote • Bangalore Urban+1💼 Full-time🗓 2026-06-25

Core

Ensuring production system health, preventing incidents, and optimizing performance for Chevron's end-to-end service operations stack.

Role type

Senior Site Reliability Engineer (SRE)

Builds

Production environments and automated resolutions for service operations

Domain

IT Infrastructure, Cloud & On-prem Architecture, Cybersecurity

Deliverable

infrastructure

Required skills

Full-stack infrastructure troubleshooting, Network administration (CISCO/Juniper), Identity & Access Management (Azure AD, SAML, OpenID), On-prem & cloud architecture, Windows & Linux OS, Performance monitoring, API integration, Automation scripting (Ansible, PowerShell, KQL, Shell)

Preferred skills

Software development lifecycle knowledge, CI/CD pipeline management, Agile team experience (Scrum/Kanban)

Technologies

Cisco, Juniper, Active Directory, Azure AD, Git, GitHub, Ansible, PowerShell, KQL, Shell

Responsibilities

Baseline service performance to prevent incidents, Lead agile teams in troubleshooting system problems, Serve as technical resource during critical incidents, Facilitate SRE technical assessments and maturity planning, Develop automation scripts to eliminate toil, Oversee production environments and monitor availability, Optimize system performance and align with DevOps best practices

Seniority

Senior, hands-on IC

Rewrite
## About the Role Exp : 8-12 years Shift Timing : 13:30 - 22:30 WFO: Tuesday - Friday (WFH - Monday) ## Responsibilities - Utilizing broad full-stack knowledge and experience for proactive incident prevention by baselining against expected service performance, improving processes from lessons learned, and using data analytics to identify problem areas and operational gaps. - Leading product line agile teams in troubleshooting and resolving system problems, including analyzing application and critical system performance. - Serving as a technical resource during critical and major incidents supporting multiple technologies. - Facilitating SRE technical assessments, identifying gaps, and providing recommendations to product teams on SRE maturity journey plans based on Chevron's SRE framework. - Finding opportunities to avoid future issues by improving logging and creating automated resolutions based on triggers. Developing automation scripts for repetitive tasks to eliminate toil/operations support activities. - Overseeing production environments by monitoring availability and maintaining a holistic view of system health. - Measuring and optimizing system performance, continuously seeking innovation and improvement to meet customer needs. Aligning, collaborating, and building relationships with peers, company leadership, subject matter experts, and users to enhance knowledge of end-to-end DevOps/Site Reliability Engineering best practices. - Collaborating with SRE Community of Practice thought leaders to define SRE capabilities and best practices and integrate the capability framework throughout the organization. ## Required Qualifications - Hands-on experience as an IT professional with knowledge of full-stack infrastructure and experience troubleshooting incidents and production issues. - Working knowledge in several technology disciplines required for the full end-to-end service operations stack: network administration & security (CISCO/Juniper), identity & access management (Active Directory, Azure AD, SAML, OpenID Federation, certificates, and keys), cybersecurity, on-prem & cloud architecture, Windows & Linux OS, performance monitoring & management, troubleshooting (application & database), change management, API integration, and automation (Ansible, PowerShell, KQL, or Shell scripting). - Strong communication (written/verbal) and facilitation skills with the solid ability to convey business and technical information to a diverse audience. - Strong analytical and problem-solving skills with the ability to engage difficulties with persistence. ## Preferred Qualifications - Basic understanding of the software development lifecycle and software engineering best practices, including code management (Git/GitHub) and CI/CD pipeline. - Experience working in an agile team (Scrum/Kanban) is considered a plus.
Sourced via wellfound · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Wellfound ↗