CareerPlanGet AI match score →

Site Reliability Engineer- (4- 7 years) Work Location - Bangalore

Bangalore, India💼 Full-time🗓 2026-06-11 → 2026-07-31

Required skills

Bachelor’s degree in computer science or related field, Technical Expertise: 4-7 years, Solid grasp of Database Architecture and Administration which includes Designing, Provisioning, Upgrading, Operating, Backups, Security, Performance for Cassandra and Neo4J, Good hands-on experience working and administering all aspects of MS-SQL with focus on real-time troubleshooting using DMVs, Execution Plans, and Extended Events; Maintain High Availability and Disaster Recovery (HADR) via AlwaysOn and Mirroring; Automate routine operational tasks using T-SQL and PowerShell, Experience and proficiency in managing large footprints through all aspects of database lifecycle, including installations, patching, and upgrades via automation-first approach, Understand IT processes including architecture, design, implementation, and operations, Experience with CI/CD framework and tools like GIT, Jenkins, Experience with automating DB tasks using python, Ansible and AI tools

Preferred skills

Solid grasp on building scalable databases using hybrid cloud infrastructure, Experience and practice with public cloud like AWS, GCP, or Azure, Kubernetes preferred, Able to establish relationships, be culturally sensitive, have goal alignment, have learning agility, be self-motivated, able and willing to help where help is needed, Exposure to Agile development practices. Working with geographically distributed teams

Technologies

Cassandra, MS-SQL, Neo4J, T-SQL, PowerShell, GIT, Jenkins, Python, Ansible, AI tools, AWS, GCP, Azure, Kubernetes

Responsibilities

Design, write and build tools to improve the reliability, availability and scalability of our Cassandra/MS-SQL/Neo4J Cloud Database Platform Offerings, Augment existing instrumentation to build a cohesive picture of the characteristics of our systems with special attention to points of failure, Design and develop improvements, focused on resilience, to our production systems to achieve and surpass SLOs, Help improve our operational practices to minimize service disruptions, Work with our Service Assurance team to modernize and improve our monitoring and alerting stack, Design new tools to monitor alerts that help discover failures or issues before our customers, Work with engineers to identify root cause and fix issues. Influence, design and create new architectures, standards and methods for large-scale enterprise systems, Maintain services once they are live by measuring and monitoring availability, latency and overall system health, Work with AI tools/agents to maintain highly available environments and optimize operations

Seniority

4-7 years

Domain

Cloud computing, Database administration, DevOps, Agile methodology

Sourced via workday · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Workday ↗