Talent Apply
Log in
All jobs
I

Senior Manager, Site Reliability Engineering

Intuit
Mountain View, California
On-site

About this role

Senior Manager, Site Reliability Engineering

Intuit is the global financial technology platform that powers prosperity for the people and communities we serve. With approximately 100 million customers worldwide using products such as TurboTax, Credit Karma, QuickBooks, and Mailchimp, we believe that everyone should have the opportunity to prosper. We never stop working to find new, innovative ways to make that possible.

Job Overview

About the Team

Intuit's Infrastructure and Site Reliability organization owns the operational backbone that keeps QuickBooks, TurboTax, Credit Karma, and Mailchimp running for hundreds of millions of customers. The Fintech Platform Systems Engineering team builds and operates the AWS-based infrastructure, resiliency tooling, and incident response capability that underpins Intuit's money-movement and fintech services — where availability, data integrity, and trust are non-negotiable.

The Opportunity

We're hiring a Senior Manager, Site Reliability Engineering to lead a hands-on team of 10–15 systems and reliability engineers responsible for the availability, performance, and operational health of Fintech Platform services running in AWS. This leader owns the strategy and execution behind operational excellence: driving toward a 99.999% availability bar, maturing incident management practices, and building self-healing, well-instrumented infrastructure at scale.

This is a player-coach role. You will set technical direction and organizational strategy while staying close to the systems — reviewing designs, joining incident bridges, and coaching engineers through complex production issues. You'll partner closely with software engineering, product, security, and other SRE/infrastructure leaders across Intuit to raise the bar on reliability company-wide.

A defining priority for this role is AI Ops: embedding AI-driven, autonomous operations into how the team runs infrastructure. You will lead the shift from manual, human-triggered response toward self-healing systems that detect, diagnose, and remediate issues autonomously — reducing developer toil, cutting MTTR, and freeing engineering capacity to focus on higher-value work. Done well, this delivers 3x the operational impact of the team today and directly accelerates the pace at which we deliver value to customers.

Responsibilities

  • Own end-to-end operational excellence for Fintech Platform services: define and drive the strategy for achieving and sustaining 99.999% availability across customer-facing and internal systems.

  • Lead, grow, and directly manage a team of 10–15 systems/site reliability engineers — hiring, mentoring, setting goals, and developing the next generation of technical leaders.

  • Act as a hands-on technical leader: participate in architecture and design reviews, write and review code/IaC

Your next opportunity starts here

Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.

Download app

Or sponsor Premium for someone who's job hunting →