System Development Engineer, Elastic Disaster Recovery, AWS Elastic Disaster Recovery
Core
Build automation, tooling, and operational infrastructure for AWS Elastic Disaster Recovery to ensure reliable, secure, and efficient replication and recovery across heterogeneous environments.
Role type
Senior Systems Development Engineer (Infrastructure Automation)
Builds
Automation for infrastructure provisioning, deployments, and operational workflows for AWS DRS fleet
Domain
Cloud Infrastructure / Disaster Recovery / Multi-OS & Multi-Architecture
Deliverable
production ML models | product features | infrastructure
Required skills
Python, Ruby, Golang, Java, C++, C#, Rust, Linux/Unix, CI/CD pipelines, distributed systems at scale
Preferred skills
Cross-platform tooling development, self-healing system design, capacity management, root-cause analysis
Technologies
AWS, Linux distributions, Windows, x86/64, ARM64 (Graviton)
Responsibilities
Design and build software for infrastructure provisioning and deployment automation; improve CI/CD pipelines and rollback mechanisms; develop tooling for cross-platform OS and architecture support; implement monitoring and self-healing systems; tune systems for scaling and performance; drive down incident volume through programmatic fixes; partner with security teams to harden the service.
Seniority
Senior, hands-on IC