Staff Infrastructure Engineer

Posted 2 Days Ago
Be an Early Applicant
New York, NY, USA
In-Office
220K-270K Annually
Senior level
Payments • Software • Automation
Building the AI Operating System for Revenue.
The Role
Own AWS infrastructure, platform evolution, infrastructure as code, container foundations, CI/CD, observability, and reliability practices. Lead incident response, define SLIs and SLOs, improve developer environments, design resilient distributed systems, automate operational work, mentor engineers, and influence infrastructure decisions across the organization.
Summary Generated by Built In

Tabs is the AI Operating System for Revenue, built for modern finance and accounting teams. It combines deep revenue and accounting expertise with the agents and applications needed to run revenue work end to end. Tabs understands customer and contract context, applies accounting logic, and executes critical workflows with built in controls, auditability, and human oversight. With Tabs, finance teams can move from manually managing revenue workflows to directing outcomes while the system executes the work.

About the Role

We're looking for a Staff Infrastructure & Reliability Engineer to own the foundation Tabs runs on: our AWS environment, how we ship software, and how we know when something is wrong. You'll set infrastructure direction for the company as a hands-on individual contributor, partnering with our platform team, our product engineers, and the product teams who build on what you build.

Tabs is building, not maintaining. We're at the point where infrastructure is becoming a real investment area, and the decisions made in this seat will shape how the whole engineering org ships for years. Payments, billing, and revenue for high-growth companies come with real compliance requirements. The goal is to build it correctly, keep it easy to maintain, and evolve it as we grow.

This is not a corner seat. You'll be expected to shape engineering and product decisions, and you'll have engineering leadership that understands infrastructure work and will pressure-test your calls. You won't be working alone: you'll own the outcomes, with high-caliber engineers around you.

What You'll Own
  • AWS infrastructure direction and platform evolution, including the migration from ECS/Fargate toward a more modern, scalable runtime

  • Infrastructure as code and container foundations, with Terraform and Docker at the core

  • CI/CD systems with a strong emphasis on developer experience, safety, and automation (GitHub Actions today; maturing CD tomorrow)

  • Ephemeral environments and preview deploys to speed iteration and increase confidence in changes

  • Observability standards across metrics, logs, and tracing, including alert hygiene, dashboards, and SLO development

  • Incident response, on-call, postmortems, and the reliability culture that surrounds them

What You'll Do
  • Define and evolve reliability standards, SLIs, SLOs, and error budgets

  • Improve observability, alerting, and incident processes across services

  • Lead high-severity incidents hands-on and drive clear, actionable follow-ups

  • Partner with engineering teams to design resilient, scalable systems

  • Write production-quality code and automation to reduce toil and lower operational risk, so repeated problems get solved once

  • Mentor engineers and influence best practices across teams

Who You Are
  • You're a software engineer first, and your infrastructure expertise is built on that foundation

  • You've run production systems on AWS and can lead platform-level change

  • You think in systems: risk, rollback strategy, blast radius, and feedback loops

  • You treat CI/CD and environments as products that should be fast, reliable, and self-serve

  • You dig into logs and data yourself when something breaks, especially under pressure

  • You influence through trust and clarity rather than control

  • You balance pragmatism with long-term system health

  • You value learning from failure and improving processes over assigning blame

  • You communicate clearly and work well across teams

Experience
  • 8+ years in software engineering, infrastructure, or SRE roles

  • Experience in one or more modern languages such as TypeScript with a track record of writing production-quality scripts, tools, and services, and still hands-on today

  • Deep hands-on experience running production workloads on AWS, including container platforms such as ECS/Fargate

  • Expertise with infrastructure as code using Terraform, and ownership of Docker and Git workflows in production

  • Solid working knowledge of Kubernetes and Helm

  • Experience designing and running CI/CD systems such as GitHub Actions, including build parallelization and developer experience improvements

  • Deep experience with observability tooling across metrics, logs, tracing, and alerting, including defining SLIs, SLOs, and error budgets

  • Expertise operating distributed systems in production at scale, with an implementation-level understanding of messaging systems, partitioning, deploy strategies, and failure modes

  • A track record of leading high-severity incidents, debugging live production issues from logs and data, and running blameless postmortems

  • Experience proposing and evaluating multiple architectures, making trade-offs that fit the company's stage, and driving infrastructure decisions across teams

  • Experience across more than one architecture or company environment, ideally including both larger companies and high-growth startups

  • Comfortable navigating ambiguity and setting direction in a fast-moving environment

  • Experience mentoring engineers and shaping infrastructure practices across an engineering org

Nice to Have
  • Experience owning broad infrastructure surface area at a high-growth startup, including as the primary infrastructure or SRE owner

  • Experience operating Kafka or a similar distributed messaging system at scale

  • Experience building developer tooling that engineers adopt and rely on

  • Prisma expertise

This role is based onsite in our Soho office in New York City

Perks and Benefits (Full-time Employees)
  • Competitive compensation and equity

  • Unlimited PTO

  • Up to 100% employer covered monthly healthcare premium (medical, dental, vision)

  • Lunch provided via Sharebite, plus dinner for any later in office days.

  • Parental leave up to 12 weeks

  • Tax free commuter and parking benefits

  • Voluntary insurances (Life, Hospital, Critical Illness, Accident)

  • Employee Assistance Program (Rightway)

  • Free One Medical Membership

  • 401k

Tabs is an equal opportunity employer. We welcome teammates of all identities and do not discriminate on the basis of race, ethnicity, religion, gender identity, sexual orientation, age, disability, veteran status, or any other protected characteristic. We’re committed to creating an environment where everyone can grow, contribute, and feel comfortable being themselves.

Skills Required

  • 8+ years of experience in software engineering, infrastructure, or SRE roles
  • Experience with a modern programming language such as TypeScript and writing production-quality scripts, tools, and services
  • Deep hands-on experience running production workloads on AWS, including ECS or Fargate
  • Expertise with infrastructure as code using Terraform
  • Production ownership of Docker and Git workflows
  • Working knowledge of Kubernetes and Helm
  • Experience designing and running CI/CD systems such as GitHub Actions
  • Experience with build parallelization and developer experience improvements
  • Deep experience with observability across metrics, logs, tracing, and alerting
  • Experience defining SLIs, SLOs, and error budgets
  • Expertise operating distributed systems in production at scale
  • Implementation-level understanding of messaging systems, partitioning, deployment strategies, and failure modes
  • Track record of leading high-severity incidents, debugging production issues, and running blameless postmortems
  • Experience evaluating architectures, making trade-offs, and driving infrastructure decisions across teams
  • Experience across multiple architecture or company environments, ideally including larger companies and high-growth startups
  • Ability to navigate ambiguity and set direction in a fast-moving environment
  • Experience mentoring engineers and shaping infrastructure practices across an engineering organization
  • Experience owning broad infrastructure at a high-growth startup
  • Experience operating Kafka or a similar distributed messaging system at scale
  • Experience building developer tooling adopted by engineers
  • Prisma expertise
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, New York
195 Employees
Year Founded: 2023

What We Do

Finance software transformed how companies manage spend years ago. Revenue never got the same treatment. Billing, collections, and revenue recognition stayed manual, because the accounting complexity and the risk of getting it wrong made the work hard to automate. So finance teams scaled with headcount instead. As revenue grew, so did the exceptions, judgment calls, and spreadsheet work needed to support it. Meanwhile, revenue itself got more complicated. Companies now price with usage, credits, commitments, consumption, outcomes, and hybrids of them. Products change, pricing changes, and sometimes the unit of value itself changes. Static systems were never designed for that variability. We started Tabs in 2023 to fix this with AI. Today we run the contract-to-cash lifecycle as one connected system: contracts, billing, collections, payments, revenue recognition, and reporting.

Why Work With Us

Founded by operators who’ve spent decades scaling companies and backed by Lightspeed, General Catalyst, and Primary, you will have the opportunity to join a team that is rethinking how revenue runs from contract to cash.

Similar Jobs

Gusto Logo Gusto

Staff Software Engineer

Fintech • HR Tech
Easy Apply
Hybrid
3 Locations
4405 Employees
163K-247K Annually

Hercules Logo Hercules

Staff Software Engineer

Hardware • Information Technology
Remote or Hybrid
2 Locations
5116 Employees
100K-300K Annually

Faire Logo Faire

Infrastructure Engineer

eCommerce • Fintech • Machine Learning • Retail
In-Office
2 Locations
1200 Employees
247K-339K Annually

General Intuition & Medal Logo General Intuition & Medal

Staff Software Engineer

Artificial Intelligence • Machine Learning • Robotics • Generative AI
In-Office
New York City, NY, USA
41 Employees
180K-275K Annually

Similar Companies Hiring

Ford Energy Thumbnail
Automotive • Software • Energy • Utilities • Manufacturing • Renewable Energy
US
55 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
70 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account