Sustaining Operations Engineer
Core
Final point of escalation for operational troubleshooting and resolution of complex issues in Linux-based software-defined infrastructure, covering bare metal, virtualization, containerization, storage, networking, and orchestration.
Role type
Senior IC sustaining operations engineer (Linux/OpenStack/Kubernetes)
Builds
Fixes, workarounds, and upstream patches for the Ubuntu, OpenStack, Ceph, and Kubernetes ecosystem
Domain
Cloud infrastructure, Open Source, Linux Systems
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Advanced Linux troubleshooting, Deep debugging (gdb, pdb, tcpdump), Python/Go/C/C++ programming, Git source code management, Distributed systems knowledge
Preferred skills
Experience with LXD, OpenStack, Ceph, Kubernetes, QEMU/KVM, LXC/LXD, Postgresql, Mongo, Debian packaging
Technologies
Linux, KVM, Docker, LXC, LXD, Ceph, OVS, OVN, OpenStack, Kubernetes, Python, Go, C, C++, gdb, pdb, tcpdump, Git
Responsibilities
Resolve complex customer problems related to Ubuntu, OpenStack, Ceph, and Kubernetes; Debug issues and propose workarounds; Liaise with software engineers to produce patches; Participate in upstream communities; Provide subject matter expertise as final escalation point; Manage time effectively against priorities
Seniority
Senior, hands-on IC