CareerPlanGet AI match score →

Systems Development Engineer, Amazon Core Search

Tokyo, Japan💼 Full-time🗓 2026-04-14 → 2026-07-31

Core

Operate and improve the reliability of Amazon's retail search infrastructure, mitigating high-severity events and automating operational workflows.

Role type

Senior Systems Development Engineer (SRE/DevOps)

Builds

Resilient search front-end services and automated incident response systems for Amazon retail websites

Domain

E-commerce / Distributed Systems / Cloud Infrastructure

Deliverable

infrastructure

Required skills

Infrastructure automation, Linux/Unix administration, Modern programming (Python, Ruby, Golang, Java, C++, C#, Rust), Incident response, Root cause analysis, SLO management, CI/CD pipeline development

Preferred skills

CI/CD pipeline build processes

Responsibilities

Mitigate and resolve high severity production events, Own post-incident analysis processes, Automate incident response workflows, Coordinate on-call rotations across global teams, Develop innovative software solutions to reduce operational burden

Seniority

Senior, hands-on IC

Rewrite
## Responsibilities - Operational Excellence: Protect the Search customer experience and eliminate operational load. Mitigate, resolve high severity events and prevent their recurrence by owning the post-incident analysis process. Have a well established OE program to identify and drive resolution of cross-service and inter-service issues, review operational metrics, and investigate Service Level Objective (SLO) breaches. - Systems Development: Develop and own innovative software solutions. Automate incident response workflows, assist engineers on call with root-causing of high severity events, and present mitigation options. Constantly automate processes through our operational excellence program, pager load budgeting system, and top root causes evaluation. - On-Call Responsibility: Operate a follow the sun on-call rotation with a frequency of one week on-call in every four to five weeks. The rotation is daytime only with handoffs between the Tokyo (Japan), and US west coast teams. The on-call represents Search during large scale events in production, leads and co-ordinates calls to quickly mitigate customer impact. - Career Growth: Care about your career aspirations. Facilitate your growth through an increase in scope of the projects you work on over time, and that can also include working on projects for partner teams to experience new, and interesting challenges. ## Requirements - Experience in automating, deploying, and supporting infrastructure - Experience programming with at least one modern language such as Python, Ruby, Golang, Java, C++, C#, Rust - Experience with Linux/Unix ## Nice to Have - Experience with CI/CD pipelines build processes ## Benefits - We are strong advocates of a healthy work-life harmony, and continually focus on ensuring happiness within our team. - We use an Operational Excellence (OE) program to continuously improve our processes. - We partner with teams across Amazon Search to achieve highly-resilient systems, share knowledge, and automate away operational burden.
Sourced via amazon · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply at Amazon ↗