Senior SW Engineer - SRE, Production Support, Linux, Networking, CI/CD
Core
Build robust technical support and automation to maximize system availability, investigate and resolve live incidents, and ensure the security, availability, and performance of the Visa Cloud platform.
Role type
Senior Site Reliability Engineer (SRE)
Builds
Visa Cloud Platform services (IaaS/PaaS/Container as a service)
Domain
Payments technology / Cloud Infrastructure
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Linux and Windows systems administration, distributed computing environments, networking technologies (DNS, TCP/IP, SSL, Load Balancing), virtualization (Hyper-V, VMware, scvmm), CI/CD/release support, container technologies (Docker, Kubernetes), alerts and monitoring (Grafana/Prometheus), scripting (Shell, PowerShell), relational and non-relational databases (MySQL, NoSQL), full application stack management, ITIL disciplines (Event, Incident, Problem, Change)
Preferred skills
Documentation practices, proactive problem-solving attitude
Technologies
Linux, Windows, Docker, Kubernetes, Grafana, Prometheus, MySQL, Hyper-V, VMware, scvmm, Shell, PowerShell
Responsibilities
Identify and support platform and user operations for Visa Cloud services; solve platform user and reliability issues; apply advanced troubleshooting to determine root causes; partner with software and systems engineers to ensure service stability and performance; collaborate on managing end-to-end availability and performance of mission-critical services
Seniority
Senior, hands-on IC