Senior Site Reliability Engineer, NetBox Delivery

Posted Yesterday
12 Locations
In-Office or Remote
75K-195K Annually
Senior level
Cloud • Software
NetBox Labs makes it easier to build, run, and govern complex networks and infrastructure, for both humans and AI agents
The Role
Own NetBox’s build and release pipeline from image creation through Cloud and Enterprise deployment. Improve Django and PostgreSQL performance, establish observability and SLOs, strengthen software supply chain security, and support SOC 2 compliance. Participate in on-call rotations, incident response, and postmortems. Drive cross-team migrations and release processes while contributing fixes to NetBox Core when reliability issues originate in the application.
Summary Generated by Built In

NetBox Labs is seeking a Senior Site Reliability Engineer for NetBox Delivery, a new team in our Applications group.

NetBox is a product that reaches people in a few ways: our open source community runs NetBox OSS; commercial customers use NetBox Cloud (SaaS) or NetBox Enterprise (self-managed). NetBox Delivery owns everything between a NetBox Core release and a healthy, running instance on Cloud and Enterprise. We ship the software, keep an eye on it in production, and when something breaks, we fix it at the source rather than working around it. You will be one of the first engineers on this team and help shape how it works.

In this role you will:

  • Own the NetBox build and release pipeline, from base images to downstream availability on Cloud and Enterprise

  • Build the release handoff between NetBox Core and the Cloud and Enterprise teams, so new releases reach customers quickly and predictably

  • Make NetBox faster and more reliable in production, from application startup to Postgres performance

  • Build real observability for the application and the release pipeline, including monitoring, alerting, and SLOs

  • Act as the escalation point for performance and reliability issues and take fixes back to NetBox Core when the cause is in the code

  • Strengthen supply chain security and support SOC 2 compliance for the build pipeline

  • Share on-call duties and lead incident response and postmortems for your area

Requirements:
  • 5+ years in software engineering, platform engineering, or SRE, with proven experience writing robust, maintainable code

  • Production experience with Django and Postgres at scale, including schema design, migration risk, and query performance under real load

  • Strong container build skills, including base image design, Python dependency management, and supply chain security practices like vulnerability scanning and image signing

  • Hands-on experience with our stack or something close to it: AWS (EC2, VPC, IAM, RDS), Kubernetes and Helm, GitHub Actions, ArgoCD or FluxCD, Terraform, and Prometheus and Grafana

  • Hands-on experience building inside an AI-augmented development harness, including Claude Code and the workflows that make agentic tooling reliable

  • A track record of driving work across team boundaries, from writing the RFC to getting a cross-team migration done

Nice to haves:
  • Familiarity with the NetBox ecosystem or network automation

  • Open source contributions or project involvement

  • Experience working in a B2B software startup or high-growth organization

  • Deep experience with supply chain security tooling such as cosign, Sigstore, or SLSA

  • Experience operating high-throughput or performance-sensitive systems for large enterprise customers


About NetBox Labs:

NetBox Labs helps companies build and manage complex networks. We help customers accelerate network automation by delivering open, composable products and supporting the network automation community.

NetBox Labs is the commercial steward of open source NetBox, the world’s most popular network source of truth, and Orb, the next-generation open source network observability platform. Our products include NetBox Enterprise, a fully supported self-managed NetBox with advanced features, and NetBox Cloud, a secure, scalable, and reliable SaaS edition of NetBox.

NetBox powers thousands of companies, and NetBox Labs is backed by investment from Notable Capital (formerly GGV), Grafana Labs CEO Raj Dutt, Flybridge, IBM, Salesforce Ventures, and Mango Capital.

Our culture and values:
  • We own and solve problems with high attention to detail.

  • Our open source contributors, users, customers & team are all part of our community. When our community wins, we win.

  • We prioritize simplicity and think twice before adding complexity

  • Clear communication helps keep our team aligned and collaborating smoothly.

NetBox Labs is proud to be an equal opportunity employer. We believe diverse teams build better software, and we welcome applicants of every race, color, religion, gender identity, sexual orientation, national origin, age, disability, and veteran status. If you need accommodation at any point in the process, just let us know.

 

Skills Required

  • 5+ years of experience in software engineering, platform engineering, or site reliability engineering
  • Proven experience writing robust, maintainable code
  • Production experience with Django and PostgreSQL at scale
  • Experience with schema design, migration risk, and query performance under real load
  • Strong container build skills, including base image design and Python dependency management
  • Experience with software supply chain security practices, including vulnerability scanning and image signing
  • Hands-on experience with AWS EC2, VPC, IAM, and RDS, or comparable technologies
  • Hands-on experience with Kubernetes and Helm
  • Hands-on experience with GitHub Actions
  • Hands-on experience with ArgoCD or FluxCD
  • Hands-on experience with Terraform
  • Hands-on experience with Prometheus and Grafana
  • Experience working inside an AI-augmented development harness, including Claude Code and reliable agentic workflows
  • Track record of driving cross-team work from RFC creation through migration completion
  • Familiarity with the NetBox ecosystem or network automation
  • Open source contributions or project involvement
  • Experience in a B2B software startup or high-growth organization
  • Deep experience with supply chain security tools such as cosign, Sigstore, or SLSA
  • Experience operating high-throughput or performance-sensitive systems for large enterprise customers

What the Team is Saying

Natalia Kepets
Chris Veith
Natalie Pastrof
Riley Lewis
Katorey Shinault
Peter Armstrong
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
125 Employees
Year Founded: 2023

What We Do

NetBox Labs simplifies the full infrastructure lifecycle for the world’s most demanding technical environments. As the commercial steward of NetBox, the open-source infrastructure system of record trusted by 10,000+ organizations for more than a decade, NetBox Labs streamlines infrastructure procurement, modeling, deployment and management for both humans and agents. The company’s infrastructure intelligence platform powers business-critical systems at companies like ARM, CoreWeave, J.P. Morgan, Kaiser Permanente, and Riot Games that trust NetBox Labs to manage the networks and infrastructure critical to their business. Headquartered in New York City, NetBox Labs is backed by NGP, Notable Capital, Flybridge, IBM, Salesforce, and Two Sigma. NetBox Cloud, by NetBox Labs, is a fully supported, hosted solution with specific performance and service SLAs and commercial support. NetBox Cloud eliminates the administrative overhead associated with hosting and managing NetBox instances

Why Work With Us

Working at NetBox Labs means stepping on the accelerator. We ship fast, we take ownership, and we swing big. "Impact" is what guides our thoughts and actions. If you want to change the infrastructure industry and add a rocketship to your career, there’s no better place.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery

NetBox Labs Offices

Remote Workspace

Employees work remotely.

Work from almost anywhere — our team is fully remote.

Typical time on-site:
US

Similar Jobs

In-Office or Remote
12 Locations
125 Employees
85K-195K Annually
In-Office or Remote
12 Locations
125 Employees
85K-195K Annually
In-Office or Remote
12 Locations
125 Employees
75K-185K Annually
In-Office or Remote
12 Locations
125 Employees
85K-195K Annually

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account