
Role Overview
As a Sustaining Operations Engineer, you will serve as the final point of escalation for critical operational issues within the open-source stack. You will work across the entire stack—from bare metal and virtualization (KVM, LXD) to networking (OVS, OVN) and orchestration (OpenStack, Kubernetes)—to debug, troubleshoot, and drive resolutions for enterprise customers.
Key Responsibilities
- Resolve complex technical issues related to Ubuntu, OpenStack, Ceph, and Kubernetes.
- Act as the final point of escalation for operational troubleshooting.
- Debug issues, propose workarounds, and collaborate with Software Engineers on upstream patches.
- Participate in upstream open-source communities.
- Maintain clear, technical, and concise communication with internal teams and customers.
- Participate in a regular weekend working rotation.
- Travel internationally up to 10% of the time for team sprints and conferences.
What We Are Looking For
- Professional experience troubleshooting advanced Linux issues.
- Strong background in Computer Science, STEM, or a related field.
- Deep expertise in at least one of: Linux, LXD, OpenStack, Ceph, or Kubernetes.
- Proficiency in debugging with Python, Go, C, or C++.
- Experience using diagnostic tools such as gdb, pdb, and tcpdump.
- Familiarity with git source code management.
- An exceptional academic track record.
What We Offer
- A distributed work environment with twice-yearly in-person team sprints.
- Annual compensation reviews and performance-driven bonuses.
- USD 2,000 annual personal learning and development budget.
- Comprehensive benefits including maternity/paternity leave and an Employee Assistance Programme.
- Global travel opportunities for team meetings and events.
Benefits
Open to
Worldwide
Sign in to track applications and earn points.