Suggested rewrite: Led a cross-functional initiative that improved [business outcome] by [measurable result], demonstrating experience relevant to this role...
Senior DevOps Engineer, Observability
Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.
- Salary
- USD 75k–85k / yr
- Department
- DevOps & Infrastructure
- Employment
- Full Time
- Experience
- Senior
- Published
- Apply before
- 4 Nov 2026
- Listing views
- 164
- Application actions
- 14
Make your next move.
Prepare your resume, explore your fit, and draft a cover letter for this opportunity.
The role, at a glance.
NetBox Labs is hiring a Senior DevOps Engineer focused on observability, platform engineering, and reliable cloud infrastructure. The role owns infrastructure automation, AWS and Kubernetes operations, Terraform and Helm workflows, CI/CD tooling, and monitoring, alerting, and SLO practices. This engineer will partner with product teams to improve internal developer platforms while supporting SOC 2-oriented security and compliance controls. The position requires substantial startup experience, hands-on scripting ability, strong documentation habits, and participation in an on-call rotation. Experience scaling multi-tenant observability systems or working with CDC and event-streaming technologies is particularly valuable.
Role DNA
A quick view of the complexity, pace, ownership and collaboration implied by the job description.
Pace & Pressure
5/5Autonomy Level
5/5Communication Load
4/5Salary analysis
Estimated compensation compared with the broader US market for similar roles.
Core skills
Skills and capabilities most closely associated with this opportunity.
Sample interview questions
I would describe the EKS architecture, networking, IAM model, cluster upgrade approach, workload isolation, autoscaling, and operational runbooks I implemented. I would also quantify reliability, deployment speed, cost, or incident-reduction outcomes where possible.
A strong answer should explain how metrics, logs, traces, dashboards, alert routing, and SLOs were designed together. It should include how alert noise was reduced, how service ownership was established, and how the platform improved incident detection or resolution.
I use reusable, versioned modules and charts; separate environment configuration; peer-reviewed pull requests; automated plan and validation checks; and controlled promotion workflows. I also manage state securely, minimize manual changes, and document recovery procedures for failed deployments.
I would first measure current lead time, failure rate, test duration, and developer pain points. Then I would standardize reusable GitHub Actions workflows, add caching and parallelization, implement policy and security checks, provide clear deployment visibility, and create self-service paths with safe defaults.
I would clearly cover detection, triage, mitigation, stakeholder communication, root-cause analysis, and follow-up work. The best example demonstrates calm prioritization, blameless learning, and concrete preventive actions such as improved alerts, capacity changes, automation, or runbook updates.
About this role.
NetBox Labs is seeking a Senior Devops Engineer to join our rapidly expanding engineering team. We have multiple positions at different seniority levels open across several teams.
A Senior DevOps Engineer at NetBox Labs is an infrastructure-minded engineer who treats operations like software – automating, instrumenting, and operating our systems with code. They focus on building and maintaining reliable, scalable infrastructure through IaC, owning the reliability and performance of what they deploy. While their primary expertise lies in infrastructure and operations, they’re comfortable navigating and contributing to software systems, collaborating closely with developers to bridge the gap between code and production.
This role is ideal for someone who thrives in fast-paced environments, enjoys solving infrastructure challenges at scale, and treats internal platforms like products – with velocity in mind.
You’ll play a key role in designing, building, and operating the infrastructure and tooling that power NetBox Labs products – from infrastructure automation to CI/CD pipelines to observability systems, you’ll help create the foundations that our teams rely on to ship high-quality software quickly and reliably.
Responsibilities
Design, build, and maintain infrastructure systems supporting NetBox Labs’ SaaS and On-Premise engineering needs.
Operate and optimize AWS and other cloud infrastructure with a focus on cost efficiency, security, and performance.
Contribute to internal platform tooling, CI/CD automation (GitHub Actions), and developer self-service capabilities.
Enhance observability and incident response systems, including monitoring, alerting, and SLOs.
Collaborate with stream-aligned product teams to understand their needs and continuously improve internal platforms.
Help enforce and improve security and compliance standards, including SOC 2 controls.
Contribute to documentation, onboarding materials, and internal support processes.
Participate in on-call rotation
Requirements
5+ years of experience in DevOps, SRE, or platform engineering roles.
2+ years of experience at a B2B software startup.
Strong experience with AWS (EC2, VPC, IAM, RDS, etc.) especially EKS/Kubernetes and infrastructure-as-code (Terraform, Helm).
Experience with CI/CD pipelines and automation tooling—ideally GitHub Actions.
Familiarity with observability tools like Prometheus, Grafana, Mimir, Loki, or similar.
Proficiency in Python, Go, or shell scripting.
Comfortable operating in a fast-paced, ambiguous startup environment.
Strong communication and documentation skills.
Experience with Change Data Capture (CDC) and event streaming systems OR experience with scaling large, multi-tenant observability systems including ingest, analysis and alerting.
Nice to Have
Experience working in a SOC 2 compliant or security-focused environment.
Experience with MQTT, AMQP and other messaging technologies
Familiarity with the NetBox ecosystem or network automation tooling.
Open-source experience or contributions.
Experience with AI tools (e.g. Copilot, ChatGPT, Cursor)
About NetBox Labs:
NetBox Labs helps companies build and manage complex networks. We help customers accelerate network automation by delivering open, composable products and supporting the network automation community.
NetBox Labs is the commercial steward of open source NetBox, the world’s most popular network source of truth, and Orb, the next-generation open source network observability platform. Our products include NetBox Enterprise, a fully supported self-managed NetBox with advanced features, and NetBox Cloud, a secure, scalable, and reliable SaaS edition of NetBox.
NetBox powers thousands of companies, and NetBox Labs is backed by investment from Notable Capital (formerly GGV), Grafana Labs CEO Raj Dutt, Flybridge, IBM, Salesforce Ventures, and Mango Capital.
Our culture and values:
We own and solve problems with high attention to detail.
Our open source contributors, users, customers & team are all part of our community. When our community wins, we win.
We prioritize simplicity and think twice before adding complexity
Clear communication helps keep our team aligned and collaborating smoothly.
NetBox Labs is proud to be an equal opportunity employer. We believe diverse teams build better software, and we welcome applicants of every race, color, religion, gender identity, sexual orientation, national origin, age, disability, and veteran status. If you need accommodation at any point in the process, just let us know.
This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.
Apply now.
Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.
Continue on the employer website
Protect your personal information and never pay to secure an interview or job offer. .
