← All jobs
Staff Infrastructure Engineer
New York City, NY · On-site · FullTime · Engineering
$220K – $270K
Apply well, not just fast
Create a free account and upload your resume to get a match score, keyword gaps, a tailored resume, a cover letter and interview prep for this job.
About the role
AWSCI/CDObservabilityDockerTerraformGitHub ActionsSRETypeScriptKafkaKubernetesGitDistributed SystemsIncident Response
Tabs is the AI Operating System for Revenue, built for modern finance and accounting teams. It combines deep revenue and accounting expertise with the agents and applications needed to run revenue work end to end. Tabs understands customer and contract context, applies accounting logic, and executes critical workflows with built in controls, auditability, and human oversight. With Tabs, finance teams can move from manually managing revenue workflows to directing outcomes while the system executes the work.
ABOUT THE ROLE
We're looking for a Staff Infrastructure & Reliability Engineer to own the foundation Tabs runs on: our AWS environment, how we ship software, and how we know when something is wrong. You'll set infrastructure direction for the company as a hands-on individual contributor, partnering with our platform team, our product engineers, and the product teams who build on what you build.
Tabs is building, not maintaining. We're at the point where infrastructure is becoming a real investment area, and the decisions made in this seat will shape how the whole engineering org ships for years. Payments, billing, and revenue for high-growth companies come with real compliance requirements. The goal is to build it correctly, keep it easy to maintain, and evolve it as we grow.
This is not a corner seat. You'll be expected to shape engineering and product decisions, and you'll have engineering leadership that understands infrastructure work and will pressure-test your calls. You won't be working alone: you'll own the outcomes, with high-caliber engineers around you.
WHAT YOU'LL OWN
- AWS infrastructure direction and platform evolution, including the migration from ECS/Fargate toward a more modern, scalable runtime
- Infrastructure as code and container foundations, with Terraform and Docker at the core
- CI/CD systems with a strong emphasis on developer experience, safety, and automation (GitHub Actions today; maturing CD tomorrow)
- Ephemeral environments and preview deploys to speed iteration and increase confidence in changes
- Observability standards across metrics, logs, and tracing, including alert hygiene, dashboards, and SLO development
- Incident response, on-call, postmortems, and the reliability culture that surrounds them
WHAT YOU'LL DO
- Define and evolve reliability standards, SLIs, SLOs, and error budgets
- Improve observability, alerting, and incident processes across services
- Lead high-severity incidents hands-on and drive clear, actionable follow-ups
- Partner with engineering teams to design resilient, scalable systems
- Write production-quality code and automation to reduce toil and lower operational risk, so repeated problems get solved once
- Mentor engineers and influence best practices across teams
WHO YOU ARE
- You're a software engineer first, and your infrastructure expertise is built on that foundation
- You've run production systems on AWS and can lead platform-level change
- You think in systems: risk, rollback strategy, blast radius, and feedback loops
- You treat CI/CD and environments as products that should be fast, reliable, and self-serve
- You dig into logs and data yourself when something breaks, especially under pressure
- You influence through trust and clarity rather than control
- You balance pragmatism with long-term system health
- You value learning from failure and improving processes over assigning blame
- You communicate clearly and work well across teams
EXPERIENCE
- 8+ years in software engineering, infrastructure, or SRE roles
- Experience in one or more modern languages such as TypeScript with a track record of writing production-quality scripts, tools, and services, and still hands-on today
- Deep hands-on experience running production workloads on AWS, including container platforms such as ECS/Fargate
- Expertise with infrastructure as code using Terraform, and ownership of Docker and Git workflows in production
- Solid working knowledge of Kubernetes and Helm
- Experience designing and running CI/CD systems such as GitHub Actions, including build parallelization and developer experience improvements
- Deep experience with observability tooling across metrics, logs, tracing, and alerting, including defining SLIs, SLOs, and error budgets
- Expertise operating distributed systems in production at scale, with an implementation-level understanding of messaging systems, partitioning, deploy strategies, and failure modes
- A track record of leading high-severity incidents, debugging live production issues from logs and data, and running blameless postmortems
- Experience proposing and evaluating multiple architectures, making trade-offs that fit the company's stage, and driving infrastructure decisions across teams
- Experience across more than one architecture or company environment, ideally including both larger companies and high-growth startups
- Comfortable navigating ambiguity and setting direction in a fast-moving environment
- Experience mentoring engineers and shaping infrastructure practices across an engineering org
NICE TO HAVE
- Experience owning broad infrastructure surface area at a high-growth startup, including as the primary infrastructure or SRE owner
- Experience operating Kafka or a similar distributed messaging system at scale
- Experience building developer tooling that engineers adopt and rely on
- Prisma expertise
This role is based onsite in our Soho office in New York City
PERKS AND BENEFITS (FULL-TIME EMPLOYEES)
- Competitive compensation and equity
- Unlimited PTO
- Up to 100% employer covered monthly healthcare premium (medical, dental, vision)
- Lunch provided via Sharebite, plus dinner for any later in office days.
- Parental leave up to 12 weeks
- Tax free commuter and parking benefits
- Voluntary insurances (Life, Hospital, Critical Illness, Accident)
- Employee Assistance Program (Rightway)
- Free One Medical Membership
- 401k
Tabs is an equal opportunity employer. We welcome teammates of all identities and do not discriminate on the basis of race, ethnicity, religion, gender identity, sexual orientation, age, disability, veteran status, or any other protected characteristic. We’re committed to creating an environment where everyone can grow, contribute, and feel comfortable being themselves.