Infrastructure / DevOps Engineer
Core
Build and maintain the foundation of compute infrastructure for a 3D scanning company, ensuring multi-GPU clusters are reliable and deployments are reproducible.
Role type
Senior Infrastructure / DevOps Engineer
Builds
Heterogeneous CPU/GPU compute clusters, container orchestration environments, and storage solutions for 3D model scanning workloads.
Domain
High-Performance Computing (HPC) / Cloud Infrastructure / 3D Data Processing
Deliverable
infrastructure
Required skills
Linux system administration, container orchestration (Kubernetes, Docker), infrastructure as code (Terraform, Ansible), HPC cluster management (Slurm), storage system design (NFS, Ceph), networking, hardware infrastructure management (IPMI, BMC)
Preferred skills
Cloud services (AWS), bare-metal provisioning (MaaS)
Technologies
Kubernetes, Docker, Terraform, Ansible, Slurm, Prometheus, Grafana, IPMI, NFS, Ceph
Responsibilities
Provision and maintain heterogeneous compute clusters; implement dynamic compute and storage provisioning; design storage solutions at hardware and software levels; manage container orchestration systems; design and maintain infrastructure as code; build and optimize job scheduling systems; set up monitoring and observability infrastructure; profile and optimize system-level performance; manage networking and secure access; handle reliability concerns including disaster recovery.
Seniority
Senior, hands-on IC