Sustaining Operations Engineer
Core
Final point of escalation for operational troubleshooting and resolution of critical issues in Linux-based software-defined infrastructure, covering bare metal, virtualization, containerization, storage, networking, and orchestration.
Role type
Senior IC sustaining operations engineer (Linux infrastructure)
Builds
Stability and fixes for Ubuntu, OpenStack, Ceph, and Kubernetes environments serving enterprise customers
Domain
Cloud infrastructure, open source, Linux systems
Deliverable
production ML models | product features | dashboards & analysis | research | client delivery | infrastructure | physical/clinical work
Required skills
Advanced Linux troubleshooting, debugging with gdb/pdb/tcpdump, Python/Go/C/C++ programming, OpenStack, Ceph, Kubernetes, LXD/LXC, KVM/QEMU, distributed systems, git
Preferred skills
Kernel or userspace experience, Debian packaging, Postgresql, Mongo, upstream community participation
Technologies
Linux, KVM, Docker, LXC, LXD, Ceph, OVS, OVN, OpenStack, Kubernetes, Python, Go, C, C++, gdb, pdb, tcpdump, git
Responsibilities
Resolve complex customer problems related to Ubuntu, OpenStack, Ceph, and Kubernetes; Debug issues and propose workarounds; Liaise with software engineers to produce patches; Participate in upstream communities; Provide subject matter expertise as final escalation point; Manage time against priorities; Participate in team activities to improve processes and documentation
Seniority
Senior, hands-on IC