All remote jobs
Open role
Remote opportunity atOkta

Manager, Site Reliability Engineering (Auth0)

Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.

Published
14Listing views
1Application actions
29 Sep 2026Apply before
Opportunity details

About this role.

AI Summary

Okta is seeking a hands-on SRE manager to lead the Auth0 reliability organization and set its technical direction. The role combines people leadership, cross-functional roadmap execution, architecture input, and direct participation in a 24/7 on-call rotation. The manager will improve observability, automation, incident response, and resilience across cloud-native systems. Candidates need significant SRE or software-engineering leadership experience, strong AWS/Azure, Kubernetes, Terraform, and Go or Python expertise, plus the ability to lead distributed teams. U.S. Person status is required because the position may access federal environments or protected federal data.

Role DNA

A quick view of the complexity, pace, ownership and collaboration implied by the job description.

Job Complexity

5/5
EasyHard

Pace & Pressure

5/5
RelaxedFast-paced

Autonomy Level

5/5
GuidedFull ownership

Communication Load

5/5
IndependentCollaborative
AI insightThis is a senior technical leadership role responsible for reliability at authentication-platform scale while remaining active in incident response and operational engineering. It requires deep cloud-native expertise, people management, strategic influence, and calm, precise communication during high-pressure outages.

Salary analysis

Estimated compensation compared with the broader US market for similar roles.

Estimated job medianMarket rate
$216,400
US market range$190k–$270k
AI insightThe disclosed annual base-salary range is $182,000 to $250,800 USD, with a midpoint of $216,400. For a U.S.-based SRE engineering manager with substantial cloud-native, incident-management, and people-leadership responsibility, an estimated base-salary market range is $190,000 to $270,000 USD annually; equity and bonus may be additional.

Core skills

Skills and capabilities most closely associated with this opportunity.

Sample interview questions
How would you balance hands-on incident response responsibilities with managing and developing an SRE team?

I would establish clear incident roles, escalation paths, and rotation ownership so I can remain technically credible without becoming a bottleneck. Outside incidents, I would use one-on-ones, design reviews, pairing, and post-incident follow-through to coach engineers, delegate meaningful ownership, and grow technical leaders.

Describe how you would reduce operational toil in a Kubernetes-based platform.

I would start by measuring recurring manual work, paging patterns, and time spent on common failure modes. I would prioritize high-frequency and high-risk workflows for automation, improve deployment and rollback controls, codify infrastructure through Terraform, and validate results through reduced alert volume, lower mean time to recovery, and improved engineer capacity.

What metrics would you use to assess reliability for an authentication platform such as Auth0?

I would define service-level indicators around successful authentication transactions, latency, availability, error rates, and dependency health. I would pair those with operational measures such as MTTA, MTTR, incident recurrence, error-budget consumption, alert quality, and the percentage of remediation that is automated.

How do you lead a blameless post-incident review after a serious outage?

I focus the review on reconstructing facts, decisions, system behavior, and contributing conditions rather than assigning individual fault. The outcome should include prioritized corrective actions with owners and due dates, such as monitoring improvements, runbook updates, architectural changes, and resilience testing, followed by leadership tracking until completion.

How would you influence product and platform teams to adopt reliability standards?

I would make reliability requirements actionable by incorporating SLOs, observability, capacity planning, and failure-mode reviews into design and delivery processes. I would use incident data and customer impact to explain tradeoffs, collaborate on pragmatic roadmaps, and demonstrate that reliability engineering enables faster, safer product delivery.

This analysis is generated from the job description. Salary estimates, role characteristics and sample answers are guidance, not employer-provided facts.

Secure Every Identity, from AI to Human

Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence.

This is an opportunity to do career-defining work. We’re all in on this mission. If you are too, let’s talk.

The SRE Leadership Team

The SRE Leadership Team at Okta is the backbone of our platform’s reliability and operational excellence. We are a forward-thinking group of engineers and leaders who believe that great infrastructure is invisible—it just works. Our team champions a culture of continuous learning, data-driven decision-making, and blameless incident response. We work at the intersection of product engineering, architecture, and operations to ensure Auth0 remains the trusted authentication platform for millions of users worldwide. As a Manager, Site Reliability Engineer, you’ll lead this team with a focus on scalability, resilience, and empowering engineers to grow as technical leaders.

What You’ll Be Doing

  • Lead the SRE team’s technical direction, translating organizational vision into actionable roadmaps while driving complex, cross-functional initiatives across product and platform teams
  • Operate at scale through hands-on participation in 24/7 on-call rotations (follow-the-sun weekdays, shared weekends), directly troubleshooting and remediating incidents on critical systems
  • Build infrastructure resilience, designing and implementing monitoring, alerting, and automation improvements that reduce toil and elevate operational efficiency
  • Champion reliability best practices, establishing policies and cultural standards that embed observability, resilience, and software engineering rigor into all engineering efforts
  • Mentor and develop SRE talent, elevating team capabilities through pair programming, design discussions, and code reviews while fostering a culture of continuous learning
  • Represent reliability as a senior technical leader in architectural reviews and strategic planning, ensuring reliability is a core consideration in major engineering decisions

What You’ll Bring to the Role

  • 3+ years of hands-on team leadership in SRE or software engineering roles within cloud-native environments, combined with 8+ years of total industry experience
  • Deep expertise in cloud platforms (AWS, Azure) and infrastructure as code (Terraform), with proven experience managing cloud-native architectures including containers, Kubernetes, microservices, and databases
  • Strong programming skills in Go or Python, with a track record of building and maintaining production-grade tools, automation, and infrastructure solutions
  • Data-driven mindset grounded in SRE principles: blameless culture, systematic problem-solving, and the ability to apply software engineering approaches to operational challenges
  • Exceptional communication skills—both verbal and written—enabling you to drive clarity during high-pressure incidents and articulate complex concepts to diverse stakeholders
  • Proven ability to build and lead high-performing teams in globally distributed, remote-first environments with strong interpersonal and collaboration skills
  • Strategic vision and technical depth, combining leadership acumen with hands-on technical excellence and a passion for mentoring senior engineers and shaping team direction

Extra Credit

  • Experience leading reliability initiatives that directly improved system uptime and reduced incident response times at scale
  • Contributions to open-source infrastructure or observability tooling
  • Experience designing and implementing comprehensive incident response programs and runbook automation

Additional requirements:

  • This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.

P13036

Below is the annual base salary range for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York and Washington. Your actual base salary will depend on factors such as your skills, qualifications, experience, and work location. In addition, Okta offers equity (where applicable), bonus, and benefits, including health, dental and vision insurance, 401(k), flexible spending account, and paid leave (including PTO and parental leave) in accordance with our applicable plans and policies. To learn more about our Total Rewards program please visit: https://rewards.okta.com/us.

The annual base salary range for this position for candidates located in California (excluding San Francisco Bay Area), Colorado, Illinois, New York, and Washington is between:

$182,000—$250,800 USD

The Okta Experience

We are intentional about connection. Our global community, spanning over 20 offices worldwide, is united by a drive to innovate. Your journey begins with an immersive, in-person onboarding experience designed to accelerate your impact and connect you to our mission and team from day one.

Okta is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, ancestry, marital status, age, physical or mental disability, or status as a protected veteran. We also consider for employment qualified applicants with arrest and convictions records, consistent with applicable laws.

If reasonable accommodation is needed to complete any part of the job application, interview process, or onboarding please use this Form to request an accommodation.

Notice for New York City Applicants & Employees: Okta may use Automated Employment Decision Tools (AEDT), as defined by New York City Local Law 144, that use artificial intelligence, machine learning, or other automated processes to assist in our recruitment and hiring process. In accordance with NYC Local Law 144, if you are an applicant or employee residing in New York City, please click here to view our full NYC AEDT Notice.

Apply now >

This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.

Next step

Apply now.

Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.

Did you apply?Let us know, and we’ll help you track your application.

Continue on the employer website

Protect your personal information and never pay to secure an interview or job offer. View safety guidance.

Log in to save
One quick step before you apply

Create your free account, then apply.

Build a more organized job search on Jobicy and continue to the employer's application when you're ready.

  • Never lose a promising opportunitySave roles and return to them from your dashboard.
  • See your entire search at a glanceTrack applications, stages and next steps in one place.
  • Get matched with relevant remote jobsChoose the alerts and digests that work for you.
Applying is free. The employer's application opens in a new tab.
Add alert
Jobs Talent AI Tools Salaries
Menu