Staff Infrastructure Engineer - Technical Staff

Savvy
NYC Office
Hybrid

Who this role is best for

Geared toward senior infrastructure engineers comfortable with cloud platform ownership and agent-based workflows, with a focus on AWS, EKS, and OpenTofu.

Best fit for

  • Senior engineers with deep Kubernetes and cloud platform ownership experience, who have led infrastructure transitions in high-stakes environments.
    — “3+ years as the escalation point for a cluster you owned
  • Candidates who have experience migrating from managed PaaS to self-managed infrastructure and setting IaC standards.
    — “Experience migrating off a managed PaaS onto self-managed infrastructure
  • Individuals who have worked in small infrastructure teams where operational decisions directly impact business outcomes.
    — “Small team, real consequences: You've worked on an infrastructure team of five or fewer where mistakes cost money

Things to consider

  • The role demands hands-on leadership in infrastructure, not just maintenance or support.
    — “You'll own the cloud platform: cluster architecture, networking, IAM, cost
  • Candidates must be able to work closely with product engineers, not just from a ticket queue.
    — “You'll work directly with product engineers on the problems in front of them

How to stand out

  • Highlight your experience in setting IaC standards and leading infrastructure remediation projects.
    — “Audit and course-correct: Review what we've built, identify what won't hold
  • Demonstrate your ability to design and implement operational practices like SLOs and incident response.
    — “Raise our operational bar: Build out SLOs, alerting that means something
  • Showcase your work with stateful systems like Postgres and your approach to upgrades and failover.
    — “Run our stateful systems: Own Postgres, Redis, and job processing in production
  • Emphasize your hands-on infrastructure engineering experience, not just cloud tooling.
    — “You build software: A controller, a CLI, a deploy system
  • Demonstrate fluency with agentic coding tools and your perspective on their infrastructure needs.
    — “Fluent with agentic coding tools: You use Claude Code, Codex, or similar in your own work
Pace · SteadyCollaboration · HighAutonomy · HighDecision Impact · CompanyLevel · Senior

Derived from job-description analysis by Serendipath's career intelligence engine.

What success looks like

  • Setting technical direction for cloud platform
  • Building and maintaining infrastructure standards
  • Collaborating with product engineers
Typical background
Production infrastructure experienceDevOps background

Skills & requirements

Required

AWSEKSOpentofuCluster ArchitectureNetworkingIAMCost ManagementCI/CDObservabilityPostgreSQLRedisBackground Job Tiers

Preferred

AI EnablementRevops

Stack & domain

AWSEKSOpentofuProblem-solvingOpportunity-findingInfrastructureCloudDevOps

About the role

Original posting from Savvy via Ashby

ABOUT SAVVY WEALTH:

Wealth management is a $545 billion industry that still runs on manual work. 75% of advisors offer no digital communication beyond email, and most still build financial plans by hand in Excel. Savvy is reinventing what it looks like to be a financial advisor. Founder Ritik Malhotra saw the fragmentation firsthand after seeking out his own advisor, and started Savvy to give independent advisors a modern, AI-native home.

Savvy is a registered investment advisor (RIA), and we partner with experienced financial advisors who want to grow without running the back office themselves. Advisors bring their book and join Savvy, running under their own brand (or ours), and Savvy earns a percentage of the assets they manage. In return, they get a true business-in-a-box: a proprietary tech platform and client portal, an in-house marketing team that helps them grow, a world-class investment management team, and a dedicated client services team that runs day-to-day operations and support. Advisors at Savvy service up to 50% more households and save 19 hours a week.

AI runs through everything we do. On the product side, Savvy Intelligence (released April 2026) is the only AI built for wealth managers that can see a client's complete financial picture. Internally, everyone at Savvy uses Claude and is encouraged to experiment with it, backed by a dedicated AI enablement team and a RevOps org building agents in-house.

We're a Series B company hitting our stride, with roughly 150 employees and over 500% year-over-year growth, backed by $105M from Thrive Capital, Index Ventures, Canvas Ventures, and Mark Casady (former CEO of LPL Financial). Come help us scale!

Recognition:

We're a Certified Great Place to Work and have been honored for our culture and our growth:

  • Newsweek's America's Greatest Startup Workplaces (2026)
  • Fortune Best Workplaces in New York™ (2026)
  • Great Place to Work Certified™

THE ROLE:

Savvy is hiring our first dedicated infrastructure engineer. Our platform runs on AWS, EKS, and OpenTofu, and we've built it with product engineers who picked up infrastructure work alongside shipping features. That got us here. Going forward we want someone who has run production infrastructure for years, can look hard at the decisions we've made, and set the direction from here.

You'll own the cloud platform: cluster architecture, networking, IAM, cost, IaC standards, CI/CD, observability, and the operational side of Postgres, Redis, and our background job tiers. You'll set the deploy contract and the standards that 35+ engineers build within, and you'll be the person who decides what our infrastructure looks like in a year.

The other half of the job is enablement. We're running coding agents across the engineering org and building the infrastructure to support them, which means compute, isolation, and observability problems that don't have settled answers yet. You'll work directly with product engineers on the problems in front of them rather than from behind a ticket queue.

RESPONSIBILITIES:

  • Own the platform: Set architecture and technical direction for our AWS and EKS footprint, including cluster topology, networking, IAM boundaries, and cost.
  • Audit and course-correct: Review what we've built, identify what won't hold, and sequence the remediation against a team that has to keep shipping.
  • Set the standards: Define our OpenTofu module structure, state management, deploy contract, and the patterns every engineer works within.
  • Raise our operational bar: Build out SLOs, alerting that means something, incident response, and on-call practice.
  • Run our stateful systems: Own Postgres, Redis, and job processing in production, including upgrades, failover, and capacity.
  • Build the paved road: Ship the tooling and internal services that make the platform easy for product engineers to use correctly.
  • Support agent infrastructure: Help design and run the compute layer for our coding agents and LLM workloads.
  • Mentor: Grow the engineers on the team who've been doing infrastructure work part-time into stronger operators.

MUST HAVE:

  • Production Kubernetes ownership: 3+ years as the escalation point for a cluster you owned, including upgrades, control plane, networking, and incidents. Deploying to a cluster someone else runs isn't the same job.
  • Experience inheriting infrastructure: You've taken over someone else's platform, formed a view on what to keep and what to replace, and executed it without stopping the business.
  • Small team, real consequences: You've worked on an infrastructure team of five or fewer where mistakes cost money, customers, or compliance standing. You couldn't specialize and you didn't have anyone to escalate to.
  • You build software: A controller, a CLI, a deploy system, an internal service, something other engineers depended on. Infrastructure work here is engineering, and IaC is one of the tools.
  • Terraform or OpenTofu depth: Module design, state management, and a real answer for drift.
  • Stateful systems in production: Postgres specifically, including connection management, upgrades, and failover.
  • CI/CD and observability as a practice: Not just standing up the tools, but making the signal useful.
  • Fluent with agentic coding tools: You use Claude Code, Codex, or similar in your own work, and you have opinions about what infrastructure for LLM and agent workloads should look like.

NICE TO HAVE:

  • Experience operating Rails, Sidekiq, or another stateful monolithic application (Django, Java, and similar all count).
  • Fintech or regulated-industry background: financial services, wealth management, or compliance-constrained infrastructure.
  • Experience migrating off a managed PaaS onto self-managed infrastructure.

BENEFITS:

  • Competitive salary and equity package
  • Unlimited PTO + paid company holidays
  • Access to holistic medical, dental, and vision plans
  • Company 401(k), Commuter, and HSA/FSA plans
  • NYC office in the heart of Manhattan
  • Lunch and snacks provided in the office
  • Access to virtual mental health care (Spring Health) and health concierge (Rightway) to help you find the right care
  • Access to counseling for stress management, dependent care, nutrition, fitness, legal, and financial issues (Guardian WorkLifeMatters EAP)

Source: Savvy careers (Ashby)

Similar roles