Staff Site Reliability Engineer

Reposted 25 Days Ago
Be an Early Applicant
New York, NY, USA
In-Office
220K-260K Annually
Expert/Leader
Payments • Software • Automation
Building the AI Operating System for Revenue.
The Role
Lead platform and infrastructure direction on AWS, evolve CI/CD and ephemeral environments, set observability and SLO standards, drive incident response and postmortems, mentor engineers, and build automation to reduce operational risk.
Summary Generated by Built In

Tabs is the leading AI-native revenue platform for modern finance and accounting teams. Tabs agents automate the entire contract-to-cash lifecycle, including billing, collections, revenue recognition, and reporting, to help teams eliminate manual work and accelerate cash flow.

High-growth companies like Cursor and Statsig rely on Tabs to generate invoices directly from contracts, reconcile payments in real time, and automate ASC 606 compliance.

Founded in 2023, Tabs has raised over $91 million from Lightspeed Venture Partners, General Catalyst, and Primary. The team is headquartered in New York and brings deep expertise in finance and AI.

About the Role

We’re looking for a Staff Site Reliability Engineer to lead the evolution of Tabs’ platform as we scale. In this role, you’ll operate as a senior individual contributor, partnering closely with engineering and product teams to design, build, and operate systems that are reliable, observable, and easy to develop on.

You’ll own our infrastructure direction, shape how we ship software, and set the standard for operational excellence across the company. This is a high-impact role for someone who enjoys solving complex systems problems, influencing architecture, and raising the reliability bar without becoming a gatekeeper.

What You’ll Own

  • AWS infrastructure direction and platform evolution, including the migration from ECS/Fargate toward a more modern, scalable runtime

  • CI/CD systems with a strong emphasis on developer experience, safety, and automation (GitHub Actions today; maturing CD tomorrow)

  • Ephemeral environments and preview deploys to speed iteration and increase confidence in changes

  • Observability standards across metrics, logs, and tracing, including alert hygiene, dashboards, and SLO development

  • Incident response, postmortems, and the reliability culture that surrounds them

What You’ll Do

  • Define and evolve reliability standards, SLIs, SLOs, and error budgets

  • Improve observability, alerting, and incident processes across services

  • Lead high-severity incidents and drive clear, actionable follow-ups

  • Partner with engineering teams to design resilient, scalable systems

  • Build automation to reduce toil and lower operational risk

  • Mentor engineers and influence best practices across teams

Who You Are

  • You’ve run production systems on AWS and can lead platform-level change

  • You think in systems: risk, rollback strategy, blast radius, and feedback loops

  • You treat CI/CD and environments as products that should be fast, reliable, and self-serve

  • You influence through trust and clarity rather than control

  • You balance pragmatism with long-term system health

  • You value learning from failure and improving processes over assigning blame

  • You communicate clearly and work well across teams

Experience

  • 10+ years in SRE, infrastructure, or backend engineering roles

  • Strong software engineering experience in one or more modern languages

  • Expertise operating distributed systems in production at scale

  • Deep experience with AWS, observability tooling, and CI/CD systems

  • Comfortable navigating ambiguity and setting direction in a fast-moving environment

Additional Information

This role is based onsite in our Soho office in New York City

Perks and Benefits (Full-time Employees)
  • Competitive compensation and equity

  • Unlimited PTO

  • Up to 100% employer covered monthly healthcare premium (medical, dental, vision)

  • Lunch provided via Sharebite, plus dinner for any later in office days.

  • Parental leave up to 12 weeks

  • Tax free commuter and parking benefits

  • Voluntary insurances (Life, Hospital, Critical Illness, Accident)

  • Employee Assistance Program (Rightway)

  • Free One Medical Membership

  • 401k

Tabs is an equal opportunity employer. We welcome teammates of all identities and do not discriminate on the basis of race, ethnicity, religion, gender identity, sexual orientation, age, disability, veteran status, or any other protected characteristic. We’re committed to creating an environment where everyone can grow, contribute, and feel comfortable being themselves.

Skills Required

  • 10+ years in SRE, infrastructure, or backend engineering roles
  • Strong software engineering experience in one or more modern languages
  • Expertise operating distributed systems in production at scale
  • Deep experience with AWS, observability tooling, and CI/CD systems
  • Experience with ECS/Fargate and platform migrations toward newer runtimes
  • Experience defining SLIs, SLOs, error budgets, and incident response processes
  • Ability to lead high-severity incidents and drive actionable postmortems
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, New York
195 Employees
Year Founded: 2023

What We Do

Finance software transformed how companies manage spend years ago. Revenue never got the same treatment. Billing, collections, and revenue recognition stayed manual, because the accounting complexity and the risk of getting it wrong made the work hard to automate. So finance teams scaled with headcount instead. As revenue grew, so did the exceptions, judgment calls, and spreadsheet work needed to support it. Meanwhile, revenue itself got more complicated. Companies now price with usage, credits, commitments, consumption, outcomes, and hybrids of them. Products change, pricing changes, and sometimes the unit of value itself changes. Static systems were never designed for that variability. We started Tabs in 2023 to fix this with AI. Today we run the contract-to-cash lifecycle as one connected system: contracts, billing, collections, payments, revenue recognition, and reporting.

Why Work With Us

Founded by operators who’ve spent decades scaling companies and backed by Lightspeed, General Catalyst, and Primary, you will have the opportunity to join a team that is rethinking how revenue runs from contract to cash.

Similar Jobs

TransUnion Logo TransUnion

Site Reliability Engineer

Big Data • Fintech • Information Technology • Business Intelligence • Financial Services • Cybersecurity • Big Data Analytics
Hybrid
5 Locations
13000 Employees
113K-188K Annually

NBCUniversal Logo NBCUniversal

Site Reliability Engineer

AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Remote or Hybrid
New York, NY, USA

Anthropic Logo Anthropic

Site Reliability Engineer

Artificial Intelligence • Natural Language Processing • Generative AI
In-Office or Remote
3 Locations
2500 Employees
320K-485K Annually
Hybrid
5 Locations
6000 Employees
194K-267K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account