Systems Development Engineer , Operations Infrastructure Services
Core
Design and operate scalable monitoring infrastructure on AWS to process high-volume device telemetry and network topology data, providing actionable insights to reduce downtime in global fulfillment and sortation centers.
Role type
Senior Systems Development Engineer (Operations Infrastructure)
Builds
Real-time monitoring services, data pipelines, and AI-enhanced alerting systems for Amazon Robotics fleet.
Domain
Cloud Infrastructure / DevOps / Observability
Deliverable
production ML models | product features | infrastructure
Required skills
System design and architecture, Large-scale infrastructure automation, Modern programming languages (Python, Go, Java, C++, Rust), Linux/Unix administration, CI/CD pipeline management
Preferred skills
Distributed systems at scale, Generative AI application, Incident detection optimization
Technologies
AWS, Python, Go, Java, C++, Rust, Linux, CI/CD
Responsibilities
Design and build scalable monitoring solutions, Own systems end-to-end from build to maintenance, Partner with stakeholders to automate manual interventions, Apply AI/ML techniques to improve detection quality, Ensure system reliability through testing and continuous improvement
Seniority
Senior, hands-on IC