Suggested rewrite: Led a cross-functional initiative that improved [business outcome] by [measurable result], demonstrating experience relevant to this role...
Senior Software Engineering Manager, Managed Gateways SREs
Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.
- Remote from
- USA
- Salary
- USD 153k–218k / yr
- Department
- DevOps & Infrastructure
- Employment
- Full Time
- Experience
- Senior
- Published
- Apply before
- 5 Nov 2026
- Listing views
- 37
- Application actions
- 3
Make your next move.
Prepare your resume, explore your fit, and draft a cover letter for this opportunity.
The role, at a glance.
Kong is hiring a Senior Software Engineering Manager to establish and lead its Managed Gateways Site Reliability Engineering team in Seattle. The manager will initially remain hands-on with enterprise reliability implementations while hiring, mentoring, and setting technical and operational standards for the new team. Core responsibilities include operating highly available cloud-native distributed systems, defining SLOs and SLIs, leading incident response, and reducing toil through automation and self-service tooling. The role requires deep Kubernetes, cloud platform, observability, incident-management, and infrastructure automation expertise, with Golang and API gateway experience strongly valued.
Role DNA
A quick view of the complexity, pace, ownership and collaboration implied by the job description.
Pace & Pressure
5/5Autonomy Level
5/5Communication Load
5/5Salary analysis
Estimated compensation compared with the broader US market for similar roles.
Core skills
Skills and capabilities most closely associated with this opportunity.
Sample interview questions
I would first define the service’s reliability goals, operating model, on-call expectations, and initial capability roadmap. I would hire for complementary strengths across infrastructure, automation, observability, and incident leadership, while personally owning high-risk implementations and establishing reusable engineering practices. As the team matures, I would delegate service ownership through clear objectives, runbooks, and measurable SLOs.
I would begin with customer-critical journeys such as gateway availability, request success rate, latency, configuration propagation, and control-plane accessibility. I would translate those journeys into measurable SLIs, agree on SLO targets with Product and Engineering, and implement error budgets that guide release and operational decisions. The resulting dashboards and alerts would focus on user impact rather than only infrastructure symptoms.
I establish a clear incident commander, communication cadence, and separate investigation and mitigation workstreams. The immediate priority is restoring customer impact safely, using observability data and documented rollback or failover procedures. After recovery, I facilitate a blameless review that identifies systemic causes, assigns durable corrective actions, and validates that prevention measures are completed.
I would measure repetitive manual work, prioritize tasks with high frequency or incident risk, and automate them through infrastructure-as-code, self-service workflows, and safe operational tooling. I would also improve service ownership boundaries, alert quality, and documentation so engineers spend less time on avoidable escalations. Success would be tracked through reduced manual effort, lower alert volume, faster recovery, and improved deployment reliability.
I would introduce lightweight readiness reviews that cover capacity, observability, failure modes, security, rollback plans, and support procedures early in the development cycle. By using production data, incident learnings, and explicit SLO requirements, I would make reliability tradeoffs visible and actionable. My goal would be collaborative accountability, where teams see operational readiness as a core product quality requirement rather than a late-stage gate.
About this role.
Are you ready to unlock intelligence?
If you don’t think you meet all of the criteria below but are still interested in the job, please apply. Nobody checks every box – we’re looking for candidates that are particularly strong in a few areas, and have some interest and capabilities in others.
Location:Washington
The Mission:
As a Senior Software Engineering Manager for Managed Gateways SREs, you will lead the charge in defining the reliability, scalability, and operational excellence of Kong’s critical managed services. Your team’s unwavering commitment to 99.99%+ uptime and seamless performance directly empowers developers globally, fueling the Agentic Era by providing the robust infrastructure that underpins modern API-driven applications. You’ll build this team in Toronto from the ground up, and in the early stages, you’ll stay hands-on – directly in the work, not just directing it from a distance.
What You’ll Do:
Build Kong’s Managed Gateways SRE team in Seattle, Washington from the ground up – hiring, setting the bar, and staying hands-on in enterprise implementations as the team ramps.
Act as a direct contributor to critical implementations and reliability work in the early stages, with outcomes that materially move the needle on Managed Gateways’ growth and business performance.
Lead, mentor, and grow a high-performing team of Site Reliability Engineers dedicated to Kong’s Managed Gateway offerings, in direct support of our critical enterprise customer base across the Americas and Europe.
Architect and implement robust, scalable, and fault-tolerant cloud-native systems using technologies like Kubernetes, Golang, and major cloud providers.
Own the end-to-end operational lifecycle, from proactive monitoring and alerting to incident response and blameless post-mortems, ensuring continuous service improvement.
Drive developer delight and operational efficiency through automation, self-service tooling, and streamlined workflows for deploying and managing API gateways, while proactively preventing technical debt and reducing operational toil.
Define, track, and report on key SLOs and SLIs, and advocate for architectural best practices that keep Managed Gateways performant and resilient as it scales.
Collaborate cross-functionally with Product, engineering, and support teams to influence roadmap decisions and ensure operational readiness for new features.
What You’ll Bring:
The Toolkit
Proven experience leading and managing Site Reliability Engineering or DevOps teams in a fast-paced, high-growth environment.
Deep expertise in designing, deploying, and operating highly available distributed systems on cloud platforms (AWS, Azure, or GCP).
Extensive hands-on experience with Kubernetes (k8s) and container orchestration in production environments.
Proficiency in Golang or similar modern programming languages for infrastructure automation and service development.
Strong understanding of observability principles and experience with tools such as Prometheus, Grafana, OpenTelemetry, or similar.
Demonstrated ability to manage critical incidents, perform root cause analysis, and implement effective preventative measures.
Familiarity with API gateway technologies, service mesh, or network proxies is a significant plus.
The Kong DNA
Ownership: You take full responsibility for the reliability and performance of systems, driving initiatives from conception to completion.
Urgency: You thrive in dynamic environments, prioritizing critical issues and delivering impactful solutions with speed and precision.
Collaboration: You build strong relationships, foster open communication, and work effectively across global teams to achieve shared goals.
Bonus Points:
Experience with open-source contributions or active participation in SRE/cloud-native communities.
Certifications in cloud platforms (e.g., AWS Certified DevOps Engineer) or Kubernetes (e.g., CKA, CKAD).
Background in companies focused on developer tools, infrastructure software, or API management.
#LI-KC1
About Kong:
Kong Inc., the AI Connectivity Company, is building the connectivity layer of AI. Trusted by the Fortune 500® and AI-native startups alike, Kong’s unified API and AI platform enables organizations to secure, manage, accelerate, govern, and monetize the flow of intelligence across APIs and AI traffic — on any model, any cloud. For more information, visit www.konghq.com.
This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.
Apply now.
Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.
Continue on the employer website
Protect your personal information and never pay to secure an interview or job offer. .
