All remote jobs

Senior DevOps Engineer, Observability

Review the role, location requirements, compensation details, and application process before deciding whether this opportunity fits your next career move.

Remote from
UK, USA, LATAM
Salary
USD 75k–85k / yr
Employment
Full Time
Experience
Senior
Published
Apply before
4 Nov 2026
Listing views
21
Application actions
2
Application toolkit

Make your next move.

Prepare your resume, explore your fit, and draft a cover letter for this opportunity.

AI Summary

The role, at a glance.

NetBox Labs is hiring a Senior DevOps Engineer focused on observability, platform engineering, and reliable cloud infrastructure. The role owns infrastructure automation, AWS and Kubernetes operations, Terraform and Helm workflows, CI/CD tooling, and monitoring, alerting, and SLO practices. This engineer will partner with product teams to improve internal developer platforms while supporting SOC 2-oriented security and compliance controls. The position requires substantial startup experience, hands-on scripting ability, strong documentation habits, and participation in an on-call rotation. Experience scaling multi-tenant observability systems or working with CDC and event-streaming technologies is particularly valuable.

Role DNA

A quick view of the complexity, pace, ownership and collaboration implied by the job description.

Job Complexity

5/5
EasyHard

Pace & Pressure

5/5
RelaxedFast-paced

Autonomy Level

5/5
GuidedFull ownership

Communication Load

4/5
IndependentCollaborative
AI insightThis is a senior, broad-scope infrastructure role requiring deep AWS, Kubernetes, IaC, CI/CD, observability, reliability, and security knowledge. The fast-moving startup setting, on-call expectations, and ownership of production platforms add significant technical and operational complexity.

Salary analysis

Estimated compensation compared with the broader US market for similar roles.

Estimated job medianBelow market
$80,000
US market range$140k–$190k
AI insightThe disclosed annual USD salary range is $75,000–$85,000, with a midpoint of $80,000. For a US-market Senior DevOps/Platform Engineer with AWS, Kubernetes, Terraform, CI/CD, and observability responsibilities, an estimated base-salary market range is $140,000–$190,000 annually; this market comparison is an estimate and can vary by location, company stage, and total-compensation structure.

Core skills

Skills and capabilities most closely associated with this opportunity.

Sample interview questions
How have you designed and operated Kubernetes infrastructure on AWS for production workloads?

I would describe the EKS architecture, networking, IAM model, cluster upgrade approach, workload isolation, autoscaling, and operational runbooks I implemented. I would also quantify reliability, deployment speed, cost, or incident-reduction outcomes where possible.

Describe an observability platform you built or significantly improved.

A strong answer should explain how metrics, logs, traces, dashboards, alert routing, and SLOs were designed together. It should include how alert noise was reduced, how service ownership was established, and how the platform improved incident detection or resolution.

What is your approach to managing infrastructure with Terraform and Helm across multiple environments?

I use reusable, versioned modules and charts; separate environment configuration; peer-reviewed pull requests; automated plan and validation checks; and controlled promotion workflows. I also manage state securely, minimize manual changes, and document recovery procedures for failed deployments.

How would you improve a CI/CD pipeline used by several engineering teams?

I would first measure current lead time, failure rate, test duration, and developer pain points. Then I would standardize reusable GitHub Actions workflows, add caching and parallelization, implement policy and security checks, provide clear deployment visibility, and create self-service paths with safe defaults.

Tell us about a production incident you owned while on call.

I would clearly cover detection, triage, mitigation, stakeholder communication, root-cause analysis, and follow-up work. The best example demonstrates calm prioritization, blameless learning, and concrete preventive actions such as improved alerts, capacity changes, automation, or runbook updates.

This analysis is generated from the job description. Salary estimates, role characteristics and sample answers are guidance, not employer-provided facts.
Opportunity details

About this role.

NetBox Labs is seeking a Senior Devops Engineer to join our rapidly expanding engineering team. We have multiple positions at different seniority levels open across several teams.

A Senior DevOps Engineer at NetBox Labs is an infrastructure-minded engineer who treats operations like software – automating, instrumenting, and operating our systems with code. They focus on building and maintaining reliable, scalable infrastructure through IaC, owning the reliability and performance of what they deploy. While their primary expertise lies in infrastructure and operations, they’re comfortable navigating and contributing to software systems, collaborating closely with developers to bridge the gap between code and production.

This role is ideal for someone who thrives in fast-paced environments, enjoys solving infrastructure challenges at scale, and treats internal platforms like products – with velocity in mind.

You’ll play a key role in designing, building, and operating the infrastructure and tooling that power NetBox Labs products – from infrastructure automation to CI/CD pipelines to observability systems, you’ll help create the foundations that our teams rely on to ship high-quality software quickly and reliably.

Responsibilities

  • Design, build, and maintain infrastructure systems supporting NetBox Labs’ SaaS and On-Premise engineering needs.

  • Operate and optimize AWS and other cloud infrastructure with a focus on cost efficiency, security, and performance.

  • Contribute to internal platform tooling, CI/CD automation (GitHub Actions), and developer self-service capabilities.

  • Enhance observability and incident response systems, including monitoring, alerting, and SLOs.

  • Collaborate with stream-aligned product teams to understand their needs and continuously improve internal platforms.

  • Help enforce and improve security and compliance standards, including SOC 2 controls.

  • Contribute to documentation, onboarding materials, and internal support processes.

  • Participate in on-call rotation

Requirements

  • 5+ years of experience in DevOps, SRE, or platform engineering roles.

  • 2+ years of experience at a B2B software startup.

  • Strong experience with AWS (EC2, VPC, IAM, RDS, etc.) especially EKS/Kubernetes and infrastructure-as-code (Terraform, Helm).

  • Experience with CI/CD pipelines and automation tooling—ideally GitHub Actions.

  • Familiarity with observability tools like Prometheus, Grafana, Mimir, Loki, or similar.

  • Proficiency in Python, Go, or shell scripting.

  • Comfortable operating in a fast-paced, ambiguous startup environment.

  • Strong communication and documentation skills.

  • Experience with Change Data Capture (CDC) and event streaming systems OR experience with scaling large, multi-tenant observability systems including ingest, analysis and alerting.

Nice to Have

  • Experience working in a SOC 2 compliant or security-focused environment.

  • Experience with MQTT, AMQP and other messaging technologies

  • Familiarity with the NetBox ecosystem or network automation tooling.

  • Open-source experience or contributions.

  • Experience with AI tools (e.g. Copilot, ChatGPT, Cursor)

About NetBox Labs:

NetBox Labs helps companies build and manage complex networks. We help customers accelerate network automation by delivering open, composable products and supporting the network automation community.

NetBox Labs is the commercial steward of open source NetBox, the world’s most popular network source of truth, and Orb, the next-generation open source network observability platform. Our products include NetBox Enterprise, a fully supported self-managed NetBox with advanced features, and NetBox Cloud, a secure, scalable, and reliable SaaS edition of NetBox.

NetBox powers thousands of companies, and NetBox Labs is backed by investment from Notable Capital (formerly GGV), Grafana Labs CEO Raj Dutt, Flybridge, IBM, Salesforce Ventures, and Mango Capital.

Our culture and values:

  • We own and solve problems with high attention to detail.

  • Our open source contributors, users, customers & team are all part of our community. When our community wins, we win.

  • We prioritize simplicity and think twice before adding complexity

  • Clear communication helps keep our team aligned and collaborating smoothly.

NetBox Labs is proud to be an equal opportunity employer. We believe diverse teams build better software, and we welcome applicants of every race, color, religion, gender identity, sexual orientation, national origin, age, disability, and veteran status. If you need accommodation at any point in the process, just let us know.

Apply now >

This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.

Next step

Apply now.

Follow the employer’s application method and review Jobicy’s safety guidance before sharing personal information.

Did you apply?Let us know, and we’ll help you track your application.

Continue on the employer website

Protect your personal information and never pay to secure an interview or job offer. .

Log in to save
One quick step before you apply

Sign in to continue.

Sign in or create a free account to continue to the employer's application.

Applying is free. After signing in, return to this job and select Apply Now.
Add alert
Jobs Talent AI Tools Salaries
Menu