Senior Systems Engineer
Core
Lead advanced operational, diagnostic, and engineering support for AI compute hardware in lab and data centre environments to ensure reliability and successful deployment.
Role type
Senior Systems Engineer (AI Compute Hardware)
Builds
AI compute platforms and systems infrastructure
Domain
AI hardware, Data Centre Operations, High-Performance Computing (HPC)
Deliverable
production ML models | infrastructure
Required skills
server hardware architecture knowledge, board-level debugging, failure isolation using logs/telemetry/power/thermal data, HPC/rack-scale infrastructure experience, structured root cause analysis, hardware diagnostics leadership
Preferred skills
ability to guide junior engineers, practical engineering judgement
Technologies
server blades, racks, power systems, thermal systems, network configuration, BIOS, BMC
Responsibilities
diagnose failures across server blades, racks, power systems, thermal behaviour, network configuration and BIOS/BMC issues; lead advanced operational and diagnostic support; improve hardware bring-up, validation, and troubleshooting; turn hard problems into clear root cause analysis and corrective actions; support new platforms from early bring-up through deployment
Seniority
Senior, hands-on IC