Site Reliability Engineer - Software Ops and Scaling , One Material Handling System - Software, Controls and Science
Core
Design and implement automation solutions for deploying, configuring, monitoring, and maintaining industrial control systems and material handling equipment at Amazon's scale.
Role type
Site Reliability Engineer (Software Ops and Scaling)
Builds
Scalable and reliable orchestration solutions for conveyor networks, robotic workcells, and industrial control systems across global fulfillment centers.
Domain
Industrial automation, warehouse robotics, DevOps
Deliverable
production ML models | product features | infrastructure
Required skills
Infrastructure automation, modern programming (Python, Ruby, Golang, Java, C++, C#, Rust), Linux/Unix, CI/CD pipeline development, predictive maintenance with real-time data analytics, IPC/PLC program deployment and version control
Preferred skills
CI/CD pipeline build processes
Responsibilities
Design and implement automation solutions for deployment and configuration; Develop monitoring and diagnostic tools for MHE systems; Create and maintain CI/CD pipelines for industrial control systems software; Implement predictive maintenance solutions using sensors and real-time data analytics; Automate IPC/PLC program deployment and version control across multiple facilities
Seniority
Individual Contributor