DevOps Engineer
Core
Operating and evolving highly available sovereign Google Cloud Platform (GCP) services with a focus on stability, scalability, and performance.
Role type
Senior DevOps / Site Reliability Engineer (SRE)
Builds
Sovereign GCP services under European jurisdiction
Domain
Defense, Aerospace, Cyber Security, Digital Identity, Cloud Infrastructure
Deliverable
production ML models | infrastructure
Required skills
DevOps / Site Reliability Engineering, Automation of operations and incident management, Experience in regulated industries (Banking, Insurance, Healthcare), Google Cloud Platform (GCP), Incident analysis and resolution, Post-Mortem analysis, Process standardization, Playbook development, 24x7 operations monitoring
Preferred skills
Experience in international work environments, High learning agility, Passion for modern cloud technologies
Technologies
Google Cloud Platform (GCP)
Responsibilities
Monitor SLI/SLO metrics and analyze/resolve production incidents in 24x7 operations, Collaborate with global SRE and GCP expert teams for incident analysis and service optimization, Document operational knowledge and develop automation and operational playbooks, Conduct post-mortem analyses to identify improvement potential
Seniority
Senior, hands-on IC