Lead Site Reliability Engineer

Company: Stuut
Location: San Francisco
Posted on: February 13, 2026

Job Description:

Job Description Job Description Stuut is transforming accounts receivable for B2B companies—making collections smarter and faster for companies that have historically relied on manual processes that are labor intensive and costly. Our platform is gaining traction with finance teams across industrials, chemicals, and manufacturing sectors from Fortune 10 brands to scaling midmarkets. We're backed by top-tier investors including a16z, Khosla, Activant, 1984 Ventures and Page One. The Role We’re hiring a Lead Site Reliability Engineer to drive the strategy, architecture, and execution of reliability, scalability, and operational excellence across our platform. You’ll build and scale the systems that keep Stuut highly available, performant, and resilient as we grow customers, traffic, and complexity. From defining SLOs and reliability standards to hardening infrastructure, improving observability, and guiding teams through incident response and postmortems, you’ll own the engineering rigor that allows us to ship quickly without sacrificing stability . You’ll turn strong reliability engineering into real customer trust — creating the guardrails that let product and engineering move fast with confidence. This is a hands-on technical leadership role for an engineer who excels at designing reliable distributed systems, influencing engineering practices, and leading high-impact reliability initiatives across teams. What You’ll Do Set the Reliability Strategy: define the long-term vision for site reliability, including SLOs/SLIs, error budgets, availability targets, and operational standards. Build & Scale Reliable Infrastructure: architect and maintain resilient, scalable cloud infrastructure across AWS and Kubernetes, ensuring systems are secure, fault-tolerant, and cost-effective. Own Observability & Monitoring: design and evolve monitoring, alerting, and logging systems that provide clear, actionable signals across services and environments. Lead Incident Response & Postmortems: own incident management practices, lead major incident response, and drive blameless postmortems that result in meaningful system improvements. Improve System Resilience: identify reliability risks and lead efforts around redundancy, failover, capacity planning, and graceful degradation. Optimize CI/CD & Deployment Reliability: partner with engineering teams to ensure deployments are safe, observable, and reversible; improve rollout strategies and reduce operational risk. Partner with Product & Engineering Teams: collaborate early in the development lifecycle to influence system design, scalability, and reliability tradeoffs. Reduce Toil & Improve Developer Experience: automate operational tasks, improve runbooks, and build tooling that reduces manual work and accelerates safe execution. Drive Root Cause Resolution : guide teams through deep debugging of reliability issues, ensuring fixes address underlying causes rather than symptoms. Influence Reliability Culture: promote reliability-first thinking, strong operational hygiene, and shared ownership of production systems across engineering. Mentor & Level Up the Team: coach engineers on reliability principles, incident handling, infrastructure design, and operational best practices. You Might Be a Fit If You… Have 7 years of experience in site reliability engineering, infrastructure engineering, or backend software engineering. Have designed and operated highly available, production-grade systems supporting rapid product iteration. Are fluent in Python and/or TypeScript, and comfortable building automation and tooling to support reliability goals. Have a deep experience with AWS, Kubernetes (EKS), Docker, and cloud-native architectures. Have implemented and evolved observability stacks (metrics, logs, traces) and know how to create high-signal alerting. Understand how to design, measure, and enforce SLOs, SLIs, and error budgets. Have supported systems built with modern stacks such as FastAPI, Vue.js, PostgreSQL (RDS), and event-driven architectures. Have improved reliability and operational maturity in environments using CI/CD pipelines, infrastructure as code, and modern deployment workflows. Can balance reliability, velocity, and cost — making pragmatic tradeoffs that serve customers and the business. Enjoy collaborating across Product, Backend, Frontend, and Infrastructure teams to improve system health. Thrive in a role that blends deep technical execution, system design, and leadership influence in a fast-moving environment. Compensation Top-of-market salary and equity package Benefits (for U.S.-based full-time employees) Medical, dental & vision insurance coverage for you 401(k) & Match Equity Flexible PTO Parental Leave

Keywords: Stuut, Salinas , Lead Site Reliability Engineer, IT / Software / Systems , San Francisco, California

Didn't find what you're looking for? Search again!

Let San Francisco recruiters find you. Post your resume for free!

Get San Francisco IT / Software / Systems jobs via email.

View more Salinas IT / Software / Systems jobs

Other IT / Software / Systems Jobs

Physician / Otolaryngology / California / Locums to Perm / Locum Otolaryngology(ENT) Physician job in Turlock, CA - Make $155/hr - $165/hr
Description: Aya Locums has an immediate opening for a locum Otolaryngology ENT job in Turlock, CA paying 155/hour - 165/hour. Job Details: Position: Physician Specialty: Otolaryngology ENT Start Date: 10-20-25 (more...)
Company: Aya Locums
Location: Turlock
Posted on: 02/15/2026

Board Certified Behavior Analyst, BCBA (School Based)
Description: Job Description Job Description Board Certified Behavior Analyst BCBA Join Our Mission to Empower Young Learners and Their Families About Ascend: Ascend is a multidisciplinary company dedicated to providing (more...)
Company: Ascend Rehab Services Inc
Location: Turlock
Posted on: 02/15/2026

Locum Tenens Radiation Oncologist Is Needed in California
Description: Interested in this assignment Or maybe you still have not found what you are looking for Contact one of our specialty-specific recruiters to get access to our vast network of open jobs, including some (more...)
Company: CompHealth
Location: Turlock
Posted on: 02/15/2026

Salary in Salinas, California Area | More details for Salinas, California Jobs |Salary

Travel Radiology Technician
Description: Job Description Trustaff Allied is seeking a travel Radiology Technician for a travel job in Walnut Creek, California. Job Description amp Requirements - Specialty: Radiology Technician - Discipline: (more...)
Company: Trustaff Allied
Location: Walnut Creek
Posted on: 02/15/2026

Travel Nurse RN - Emergency Room (ER) / Trauma - $1,547 per week in Turlock, CA
Description: Registered Nurse RN Emergency Room ER / Trauma Location: Turlock, CA Agency: FlexCare Pay: 1,547 per week Shift Information: Rotating - 3 days x 12 hours Contract Duration: 13 Weeks Start Date: (more...)
Company: TravelNurseSource
Location: Turlock
Posted on: 02/15/2026

Outpatient Psychiatrist
Description: At LifeStance Health, we believe in a truly healthy society where mental and physical healthcare are unified to make lives better. Our mission is to help people lead healthier, more fulfilling lives by (more...)
Company: LifeStance Health
Location: Delhi
Posted on: 02/15/2026

Occupational Therapist / OT
Description: Job Description Job Description Overview Occupational Therapist OT Home amp Community Program Make an Impact Where Life Happens Join our Rehab Without Walls team and help individuals thrive (more...)
Company: BrightSpring Health Services
Location: Walnut Creek
Posted on: 02/15/2026

Outpatient Psychiatrist
Description: At LifeStance Health, we believe in a truly healthy society where mental and physical healthcare are unified to make lives better. Our mission is to help people lead healthier, more fulfilling lives by (more...)
Company: LifeStance Health
Location: Turlock
Posted on: 02/15/2026

Outpatient Psychiatrist
Description: At LifeStance Health, we believe in a truly healthy society where mental and physical healthcare are unified to make lives better. Our mission is to help people lead healthier, more fulfilling lives by (more...)
Company: LifeStance Health
Location: Hilmar
Posted on: 02/15/2026

Outpatient Psychiatrist
Description: At LifeStance Health, we believe in a truly healthy society where mental and physical healthcare are unified to make lives better. Our mission is to help people lead healthier, more fulfilling lives by (more...)
Company: LifeStance Health
Location: Livingston
Posted on: 02/15/2026

Loading more jobs...

Lead Site Reliability Engineer

Didn't find what you're looking for? Search again!

Other IT / Software / Systems Jobs

Log In or Create An Account