Site Reliability Engineering Manager, Managed Operations
Core
Leading the launch and operations of AWS European Sovereign Cloud (ESC) for EU customers, ensuring high availability and reliability of core cloud services.
Role type
Senior Site Reliability Engineering Manager (Managed Operations)
Builds
High-availability AWS services (EC2, S3, Dynamo, Lambda, Bedrock) exclusively for EU customers
Domain
Cloud Infrastructure / Utility Computing
Deliverable
production ML models | product features | infrastructure
Required skills
Systems engineering leadership, team management, systems operations, networking, storage systems, operating systems, agile software development
Preferred skills
Full software development life cycle, continuous deployments, testing, operational excellence, AWS platforms and services
Technologies
EC2, S3, Dynamo, Lambda, Bedrock, AWS platforms
Responsibilities
Lead team execution of organizational roadmaps for ESC launch, analyze systems and software performance, cross-departmental collaboration, build coaching and mentoring teams, review key performance metrics, drive escalation calls
Seniority
Senior, hands-on IC with management responsibilities