Now hiring

Staff Infrastructure Engineer @ Tabs

USOnsiteFull-time
Apply with ResuMinder

Opens on the employer's site

About this role

Tabs is the AI Operating System for Revenue, built for modern finance and accounting teams. It combines deep revenue and accounting expertise with the agents and applications needed to run revenue work end to end. Tabs understands customer and contract context, applies accounting logic, and executes critical workflows with built in controls, auditability, and human oversight. With Tabs, finance teams can move from manually managing revenue workflows to directing outcomes while the system executes the work.

About the RoleWe're looking for a Staff Infrastructure & Reliability Engineer to own the foundation Tabs runs on: our AWS environment, how we ship software, and how we know when something is wrong. You'll set infrastructure direction for the company as a hands-on individual contributor, partnering with our platform team, our product engineers, and the product teams who build on what you build.

Tabs is building, not maintaining. We're at the point where infrastructure is becoming a real investment area, and the decisions made in this seat will shape how the whole engineering org ships for years. Payments, billing, and revenue for high-growth companies come with real compliance requirements. The goal is to build it correctly, keep it easy to maintain, and evolve it as we grow.

This is not a corner seat. You'll be expected to shape engineering and product decisions, and you'll have engineering leadership that understands infrastructure work and will pressure-test your calls. You won't be working alone: you'll own the outcomes, with high-caliber engineers around you.

What You'll OwnAWS infrastructure direction and platform evolution, including the migration from ECS/Fargate toward a more modern, scalable runtime

Infrastructure as code and container foundations, with Terraform and Docker at the core

CI/CD systems with a strong emphasis on developer experience, safety, and automation (GitHub Actions today; maturing CD tomorrow)

Ephemeral environments and preview deploys to speed iteration and increase confidence in changes

Observability standards across metrics, logs, and tracing, including alert hygiene, dashboards, and SLO development

Incident response, on-call, postmortems, and the reliability culture that surrounds them

What You'll DoDefine and evolve reliability standards, SLIs, SLOs, and error budgets

Improve observability, alerting, and incident processes across services

Lead high-severity incidents hands-on and drive clear, actionable follow-ups

Partner with engineering teams to design resilient, scalable systems

Write production-quality code and automation to reduce toil and lower operational risk, so repeated problems get solved once

Mentor engineers and influence best practices across teams

Who You AreYou're a software engineer first, and your infrastructure expertise is built on that foundation

You've run production systems on AWS and can lead platform-level change

You think in systems: risk, rollback strategy, blast radius, and feedback loops

You treat CI/CD and environments as products that should be fast, reliable, and self-serve

You dig into logs and data yourself when something breaks, especially under pressure

You influence through trust and clarity rather than control

You balance pragmatism with long-term system health

You value learning from failure and improving processes over assigning blame

You communicate clearly and work well across teams

Experience8+ years in software engineering, infrastructure, or SRE roles

Experience in one or more modern languages such as TypeScript with a track record of writing production-quality scripts, tools, and services, and still hands-on today

Deep hands-on experience running production workloads on AWS, including container platforms such as ECS/Fargate

Expertise with infrastructure as code using Terraform, and ownership of Docker and Git workflows in production

Solid working knowledge of Kubernetes and Helm

Experience designing and running CI/CD systems such as GitHub Actions, including build parallelization and developer experience improvements

Deep experience with observability tooling across metrics, logs, tracing, and alerting, including defining SLIs, SLOs, and error budgets

Expertise operating distributed systems in production at scale, with an implementation-level understanding of messaging systems, partitioning, deploy strategies, and failure modes

A track record of leading high-severity incidents, debugging live production issues from logs and data, and running blameless postmortems

Experience proposing and evaluating multiple architectures, making trade-offs that fit the company's stage, and driving infrastructure decisions across teams

Experience across more than one architecture or company environment, ideally including both larger companies and high-growth startups

Comfortable navigating ambiguity and setting direction in a fast-moving environment

Experience mentoring engineers and shaping infrastructure practices across an engineering org

Nice to HaveExperience owning broad infrastructure surface area at a high-growth startup, including as the primary infrastructure or SRE owner

Experience operating Kafka or a similar distributed messaging system at scale

Experience building developer tooling that engineers adopt and rely on

Prisma expertise

This role is based onsite in our Soho office in New York City

Perks and Benefits (Full-time Employees)Competitive compensation and equity

Unlimited PTO

Up to 100% employer covered monthly healthcare premium (medical, dental, vision)

Lunch provided via Sharebite, plus dinner for any later in office days.

Parental leave up to 12 weeks

Tax free commuter and parking benefits

Voluntary insurances (Life, Hospital, Critical Illness, Accident)

Employee Assistance Program (Rightway)

Free One Medical Membership

401k

Tabs is an equal opportunity employer. We welcome teammates of all identities and do not discriminate on the basis of race, ethnicity, religion, gender identity, sexual orientation, age, disability, veteran status, or any other protected characteristic. We’re committed to creating an environment where everyone can grow, contribute, and feel comfortable being themselves.

Skills

Engineering

Ready to apply?

Install the ResuMinder extension and we'll auto-fill the application in seconds — no rewriting.

See how your CV scores