About this role.
MaintainX is hiring a Site Reliability Engineer to improve the reliability, resilience, observability, and operational maturity of its production platform. The role partners with product and platform teams to implement SRE standards, service-health metrics, incident practices, and shared operational tooling. A central objective is enabling development teams to independently deploy, support, and operate their services. The position requires practical cloud-native and infrastructure-as-code experience, strong distributed-systems observability knowledge, and the ability to mentor and influence cross-functional engineering teams.
Role DNA
A quick view of the complexity, pace, ownership and collaboration implied by the job description.
Job Complexity
4/5Pace & Pressure
4/5Autonomy Level
4/5Communication Load
5/5Salary analysis
Estimated compensation compared with the broader CA market for similar roles.
Core skills
Skills and capabilities most closely associated with this opportunity.
Sample interview questions
I would first identify the critical user journeys and define a small number of measurable SLIs, such as availability, latency, and successful job completion. I would work with product and engineering stakeholders to set realistic initial SLOs based on historical performance and customer impact, then use the resulting error budget to guide release risk, reliability investment, and incident review decisions.
I would assess the current telemetry across logs, metrics, and traces, then establish consistent service naming, correlation IDs, dashboards, and actionable alerts tied to user-facing symptoms. I would prioritize instrumentation around critical request paths and dependencies, validate that alerts are meaningful during controlled failure scenarios, and document operational runbooks for recurring conditions.
I provide paved paths rather than creating a permanent centralized dependency: reusable deployment templates, standard dashboards, alerting defaults, runbooks, and clear ownership expectations. I pair with teams initially, offer targeted training and reviews, and measure success through reduced escalation volume, faster incident resolution, and teams independently making safe operational changes.
During an incident, I focus on clear roles, a shared timeline, frequent stakeholder updates, mitigation before root-cause certainty, and disciplined escalation. Afterward, I facilitate a blameless review that identifies contributing technical and process factors, assigns specific follow-up actions, and tracks those actions through completion to reduce recurrence.
I would connect reliability work to customer impact, engineering time lost to incidents, delivery risk, and measurable service-health data. By proposing incremental improvements that fit team roadmaps, demonstrating quick wins through shared tooling, and making reliability expectations visible in planning and release processes, I can build durable adoption without positioning SRE as a gatekeeper.
MaintainX is a leading mobile-first work execution platform for industrial and frontline teams. More than 13,000 customers, including Duracell, McDonald’s, Shell, DHL and Volvo, use MaintainX to cut unplanned downtime and run better operations, across 13.9 million managed assets and 79.5 million completed work orders.
In August 2026 MaintainX became part of Autodesk, joining Autodesk Operations Solutions, the organization unifying Autodesk’s operations platform alongside Tandem, FlexSim and Fusion Operations. Autodesk’s strategy is to converge design, make and operate into one continuous lifecycle: design an asset, build it, run it, then feed what you learn running it back into the next design. Autodesk had design and make. Operate is the phase that tells you what actually happened, and it is ours.
We’re looking for a Site Reliability Engineer to help advance MaintainX’s reliability, observability, and developer autonomy as we scale our platform.
In this role, you’ll partner closely with product and platform development teams to improve the stability, resilience, and operational readiness of our services. You’ll work alongside teams to design for reliability from the start, establish clear ownership and standards, and build shared tooling that enables teams to operate their services with confidence.
You’ll also contribute to company-wide initiatives that define how MaintainX approaches reliability software development, including observability standards, incident response practices, and service health metrics, helping the organization adopt proven industry practices at scale.
This role is well-suited for an developer who enjoys working across teams, influencing technical direction through strong development practices, and turning reliability principles into practical, scalable systems.
What You’ll Do:
Assess service maturity and provide insights to development teams
Partner with development teams to implement observability best practices
Enable development teams to become autonomous with their service deployment, support, and infrastructure
Mentor developers on reliability practices, focusing on making them self-sufficient
Act as the bridge, ear and eyes of the Platform Division teams to drive tooling and practice adoption across development teams
About You:
Deep understanding of observability practices in a distributed system environment and how it influences system design and team behaviour
Practical experience with SRE concepts (SLOs, error budgets, incident management)
3–5+ years in software development, SRE, DevOps, or production development roles with experience operating production systems
Proficient in cloud-native platforms and infrastructure-as-code concepts and tools
Working knowledge of at least one programming language (TypeScript/Node.js is a plus)
Excellent communication and collaboration abilities across technical and non-technical teams
Ability to translate complex reliability concepts into actionable guidance
You enjoy enabling teams to succeed independently and measuring success by reduced dependency on you
About Us:
MaintainX is committed to creating a diverse environment. All qualified applicants will receive consideration for employment without regard to race, colour, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status.
Our mission is to keep the physical world running. Factories, fleets, hospitals and campuses stay up because the people who maintain them have tools worth using. That is what we build.
Compensation and benefits. Base pay is one part of the package. Depending on the role, compensation may also include commission, an annual bonus and equity. Benefits differ by country. For roles in the United States, Autodesk’s benefits are described at benefits.autodesk.com. For roles in Canada and other countries, the plan differs on health coverage, retirement and leave, and your recruiter will walk you through it.
Belonging. We take pride in a culture where everyone can thrive. More at autodesk.com/company/global-belonging. More on where this is going: Autodesk CEO Andrew Anagnost on building the future of connected operations, and AOS SVP Stephen Hooper on welcoming MaintainX to Autodesk.
This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.








