AMHS Site Reliability Engineer
Core
Ensure 24x7 reliability and scalability of software controlling Samsung's automated material transport ecosystem (AMHS) and prevent fab operational disruptions.
Role type
Site Reliability Engineer (SRE) for semiconductor manufacturing execution systems
Builds
Scalable data pipelines, observability frameworks, and automation tooling for AMHS control software
Domain
Semiconductor manufacturing / Fab automation
Deliverable
production ML models | product features | dashboards & analysis | infrastructure
Required skills
Python, Java, Linux/Unix administration, back-end data engineering, ETL processes, relational/NoSQL databases, Prometheus, Grafana, RabbitMQ, Redis, Docker, Kubernetes, IPC/RPC troubleshooting
Preferred skills
SECS/GEM protocols, Manufacturing Execution Systems (MES), semiconductor manufacturing processes
Responsibilities
Ensure high availability of AMHS controlling software suite; Design and maintain scalable back-end data pipelines for fab control logs; Act as primary escalation point for software anomalies and troubleshoot network/server issues; Develop automation scripts for deployment, configuration, and system recovery
Seniority
Mid-level, hands-on IC