Infrastructure Engineer (Data Center Operations)
Core
Hands-on maintenance, provisioning, and troubleshooting of high-performance on-premise server and networking infrastructure for AI workloads.
Role type
Infrastructure Engineer (Data Center Operations)
Builds
High-performance on-premise server clusters and networking infrastructure for AI training and inference
Domain
AI hardware infrastructure / Data Center Operations
Deliverable
infrastructure
Required skills
Linux system administration, x86 server hardware, enterprise networking, physical server installation, Bash scripting, Python scripting, network troubleshooting (Layer 1-3), BIOS/firmware configuration
Preferred skills
PXE boot, NFS, RAID controllers, configuration management (Ansible), lab/R&D hardware environment experience
Technologies
Linux, Bash, Python, IPMI/iDRAC/iLO, 100G/400G networks, VLANs, static routing
Responsibilities
Physically install, rack, cable, and maintain blade servers and hardware components; Connect servers to high-speed networks and verify optics/DACs; Configure BIOS, firmware, and out-of-band management; Install and provision Linux OS and configure networking; Debug network issues at physical and OS level; Use scripting to automate routine tasks; Troubleshoot and replace failed server components with minimal downtime
Seniority
Mid-level, hands-on IC