Senior Network Engineer for Industrial AI Cloud (m/f/d)
Core
Build and automate the network platform for a sovereign industrial AI cloud hosting 10,000 GPUs, managing high-speed fabrics and core infrastructure for European manufacturers.
Role type
Senior Network Engineer (Infrastructure & Automation)
Builds
Automated network configuration, monitoring, and deployment for NVIDIA DGX B200 systems and RTX Pro Servers within a sovereign AI factory.
Domain
Industrial AI Cloud, High-Performance Computing (HPC), Data Center Networking
Deliverable
production ML models | infrastructure
Required skills
InfiniBand architecture, RoCE, Linux networking, SDN, Kubernetes, Python/Go scripting, ITIL processes, BGP/OSPF, FortiGate security management
Preferred skills
VMware Tanzu, Cumulus OS, UFM, Prometheus/Grafana, CI/CD in Kubernetes
Technologies
InfiniBand, RoCE, Cumulus OS, UFM, FortiGate, Cisco Border Gateways, NVIDIA DGX B200, RTX Pro Servers, Kubernetes, Ansible, Terraform, Helm, GitLab, GitHub Actions
Responsibilities
Provision and maintain InfiniBand switches and firewalls; develop automation scripts for network orchestration; manage OS and firmware upgrades at scale; coordinate network lifecycle activities with Data Center and IaaS/PaaS teams; mentor co-workers and act as a key technical lead.
Seniority
Senior, hands-on IC with mentorship responsibilities