Senior Manager, Support Engineering
Core
Leading planning, staffing, and execution of repair-center operations for next-generation AI infrastructure (GB200, GB300) to support routine demand and surge events across Oracle data center sites.
Role type
Senior Manager, Support Engineering (Operations)
Builds
Centralized repair-center model for AI compute platforms
Domain
AI Infrastructure / Data Center Operations
Deliverable
production ML models | infrastructure
Required skills
capacity planning, operational scaling, cross-functional coordination, staffing and resource allocation, process standardization, data-driven risk analysis, regional expansion planning, team leadership
Preferred skills
experience with GB200/GB300 platforms, establishing operating rhythms, performance tracking
Technologies
AI Support Network Oracle
Responsibilities
Lead planning, staffing, and execution of repair-center operations for AI compute platforms; own repair-center readiness including capacity planning and ramp-up; partner with Design, Architecture, Manufacturing, and Supply Chain to resolve cross-functional issues; oversee workflows for Compute Trays, CDU validation, and leak-response; establish metrics and escalation paths to maintain throughput and quality; standardize repair-center processes and best practices; support regional expansion by aligning buildouts with demand forecasts; lead hiring, onboarding, and performance management for repair-center teams; communicate site status and risks to senior leadership.
Seniority
Senior Manager, strategic execution & cross-functional leadership