Linux System & HPC Administrator (IT Infrastructure & Operations)
Core
Manage, optimize, and scale enterprise High Performance Computing (HPC) infrastructure for semiconductor design workloads.
Role type
Senior Linux System & HPC Administrator
Builds
HPC clusters, Linux environments (bare-metal, virtual, AWS), and core infrastructure services for semiconductor design
Domain
Semiconductor / High Performance Computing / IT Infrastructure
Deliverable
production ML models | infrastructure
Required skills
Linux system administration, HPC cluster management, workload management tools (LSF/Slurm), system monitoring and performance tuning, identity services and access controls, scripting (shell/Python), networking fundamentals, configuration management, automation
Preferred skills
Cloud platforms (AWS), NetApp storage, container technologies (Docker/Kubernetes), DevOps practices, CI/CD pipelines
Technologies
Red Hat Enterprise Linux, Ubuntu, VMware, Hyper-V, LSF, Slurm, Nagios, Prometheus, Grafana, Ansible, Puppet, Docker, Kubernetes, AWS, NetApp
Responsibilities
Install, configure, and maintain Linux OS and applications across large bare-metal and virtual environments; optimize HPC clusters to maximize utilization; monitor infrastructure health and resolve system issues/outages; diagnose and fix batch scheduler workload anomalies; implement security best practices and system hardening; develop automation tools to improve operational efficiency
Seniority
Senior, hands-on IC