Sustaining Operations Engineer

Remote from
🌐 Anywhere
Annual salary
Undisclosed
Salary information is not provided for this position. Check our Salary Directory to estimate the average compensation for similar roles.
Employment type
Full Time,
Job posted
Apply before
21 Aug 2026
Experience level
Senior
Views / Applies
176 / 13

About Canonical Ltd.

Trusted open source for enterprises

Actively Hiring
Verified job posting
This job post has been manually reviewed for authenticity and compliance.

AI Summary

This role is for a Sustaining Operations Engineer at Canonical, focusing on resolving complex customer issues in Linux-based software-defined infrastructure. The engineer will work across all layers of the stack, including bare metal, virtualization, containerization, storage, networking, and open-source platforms like OpenStack and Kubernetes. As the final point of escalation, the role requires deep debugging and troubleshooting skills, with responsibilities including maintaining close relationships with engineering teams and participating in upstream communities. The position is globally remote, with travel up to 10% for team events, and demands strong communication and prioritization skills. Ideal candidates have extensive experience with Linux, OpenStack, Ceph, or Kubernetes, and proficiency in debugging tools and languages like Python, Go, or C.

Role DNA

Job Complexity
Easy Hard
Pace & Pressure
Relaxed Fast-paced
Autonomy Level
Guided Full Ownership
Communication Load
Independent Highly Collaborative
AI Insight This role requires deep expertise across multiple complex open-source technologies and the ability to troubleshoot at any level of the stack, making it one of the most challenging roles.

Salary Analysis

Median Highly Competitive
$140,000
US Market
$100k – 180k
0 $198k
AI Insight No salary was provided in the job listing. Based on market data for similar senior-level sustaining engineering roles requiring expertise in Linux, OpenStack, Ceph, and Kubernetes, the estimated median salary is $140,000. The typical US market range for such positions is $100,000 to $180,000, depending on experience and location.

I am writing to express my strong interest in the Sustaining Operations Engineer role at Canonical. With extensive experience in Linux, OpenStack, Ceph, and Kubernetes, I am well-prepared to handle complex troubleshooting and drive resolution for critical customer issues. My background in debugging with tools like gdb and tcpdump, combined with my passion for open source, aligns perfectly with the responsibilities of this position. I am excited about the opportunity to work remotely and contribute to Canonical's mission of delivering exceptional enterprise support. Thank you for considering my application.

Describe your experience troubleshooting a complex issue in a Kubernetes cluster.
I once debugged a networking issue where pods were unable to communicate across nodes. Using tcpdump and Fluentd logs, I identified a misconfigured Calico policy, then corrected it and validated connectivity.
How do you handle high-pressure situations when a critical customer issue arises?
I prioritize understanding the problem thoroughly, escalate if needed, and communicate transparently with the customer. I focus on root cause analysis and provide workarounds while developing a permanent fix.
Explain a time you contributed to an upstream open-source project.
I submitted patches to the OpenStack Neutron project to fix a bug in the OVN driver. I followed the contribution guidelines and worked with maintainers to get it merged.
What is your approach to debugging a performance issue in a Ceph cluster?
I would start by checking cluster health with `ceph status`, review logs, and use tools like `ceph perf` to identify slow OSDs. Then I'd analyze disk I/O and network latency to pinpoint bottlenecks.
How do you stay current with multiple technologies like LXD, Docker, and OpenStack?
I regularly read release notes, participate in community meetings, and experiment with new features in test environments. I also follow key blogs and attend virtual conferences.

This is a fast-paced engineering role in Linux-based software-defined infrastructure and applications, covering all layers of the stack, including bare metal, virtualization (KVM) and containerization (Docker and LXC/LXD), storage (Ceph and Linux filesystems), networking (OVS, OVN and Core networking), up to OpenStack and Kubernetes, and the open source applications running on top of them.

This role is an opportunity for a technologist with a passion for Linux and open source to build a career with Canonical and drive success for our customers, community and the company. If you have great communication skills, and a passion for troubleshooting and fixing issues in technology used by millions across the world, then you will enjoy working with some of the best people in the industry at Canonical.

Location: This is a globally remote role.

This role deals with critical issues in the open source stack that require deep debugging and troubleshooting skills. Our engineers have to be able to work productively at any level of the stack above the kernel, in a wide range of applications, to understand and address the software issues at hand. Our group is critical to the success of our enterprise customers, partners and Ubuntu itself.

You will be the final point of escalation for operational troubleshooting and driving issues to resolution with workarounds, guidance, and fixes to be released upstream and in Ubuntu. 

What your day will look like

  • Resolve complex customer problems related to Ubuntu, OpenStack, Ceph and/or Kubernetes
  • Maintain a close working relationship with Canonical’s field, support and product engineering teams
  • Participate in upstream communities
  • Debug issues, propose workarounds, liaise with Software Engineers on producing a patch
  • Demonstrate good judgment in technical methods and techniques
  • Prioritize work and manage your time effectively against priorities
  • Participate in team activities to improve processes, tools, and documentation
  • Maintain clear, technical and concise communications
  • Participate in a regular weekend working rotation
  • Provide subject matter expertise as the final point of escalation on operational issues
  • Work from home and travel internationally up to 10% of work time for team meetings, events and conferences

What we are looking for in you

  • Professional experience troubleshooting advanced Linux issues
  • Background in Computer Science, STEM or similar
  • Exceptionally strong experience with either Linux, LXD, OpenStack, Ceph or Kubernetes
  • Strong debugging experience with Python, Go, C or C++ on Linux
  • Ability to troubleshoot with gdb, pdb, tcpdump or other tools
  • Familiarity with git source code repositories and branches
  • An exceptional academic track record from both high school and preferably university
  • Willingness to travel up to 4 times a year for internal events

Additional skills that you might also bring

  • You love technology and working with brilliant people
  • You are curious, flexible, articulate, and accountable
  • You value soft skills and are passionate, enterprising, thoughtful, and self-motivated
  • You have interest in, and experience with most of the following: Ubuntu Linux – kernel or userspace, Kubernetes, OpenStack, Ceph, QEMU/KVM, LXC/LXD, Python, Go, C, Postgresql, Mongo, Debian packaging, distributed systems

What we offer you

We consider geographical location, experience, and performance in shaping compensation worldwide. We revisit compensation annually (and more often for graduates and associates) to ensure we recognise outstanding performance. In addition to base pay, we offer a performance-driven annual bonus. We provide all team members with additional benefits, which reflect our values and ideals. We balance our programs to meet local needs and ensure fairness globally.

  • Distributed work environment with twice-yearly team sprints in person – we’ve been working remotely since 2004!
  • Personal learning and development budget of USD 2,000 per year
  • Annual compensation review
  • Recognition rewards
  • Annual holiday leave
  • Maternity and paternity leave
  • Employee Assistance Programme
  • Opportunity to travel to new locations to meet colleagues from your team and others
  • Priority Pass for travel and travel upgrades for long haul company events

About Canonical

Canonical is a pioneering tech firm that is at the forefront of the global move to open source. As the company that publishes Ubuntu, one of the most important open source projects and the platform for AI, IoT and the cloud, we are changing the world on a daily basis. We recruit on a global basis and set a very high standard for people joining the company. We expect excellence – in order to succeed, we need to be the best at what we do.

Canonical has been a remote-first company since its inception in 2004.​ Work at Canonical is a step into the future, and will challenge you to think differently, work smarter, learn new skills, and raise your game. Canonical provides a unique window into the world of 21st-century digital business.

Canonical is an equal opportunity employer

We are proud to foster a workplace free from discrimination. Diversity of experience, perspectives, and background create a better work environment and better products. Whatever your identity, we will give your application fair consideration.

 

#LI-Remote

 

Apply now >

This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.

How to apply

Did you apply? Let us know, and we’ll help you track your application.

See a few more

Similar Software Engineering remote jobs

Jobs Talent Salaries
Menu