Senior Infrastructure Engineer, SRE

3 hours, 46 minutes ago
Full-time
Senior
DevOps and Infrastructure
Rocket Money

Rocket Money

Rocket Money (formerly Truebill) is a top personal finance app that simplifies managing subscriptions, bills, budgets, and savings with ease, ultimately empowering users to take control of their financial lives.

Diversified Financial Services
51-250
$84M raised

Description

  • Build and improve the reliability and resiliency of production systems and services.
  • Define and review SLIs, SLOs, and error budgets for critical services and user journeys.
  • Own disaster recovery objectives, failover and restoration procedures, and regular recovery exercises.
  • Partner with product engineering teams to support service ownership and user-focused reliability metrics.
  • Evolve observability standards and platforms across metrics, tracing, and logs, including instrumentation, alerting, and cost management.
  • Strengthen incident response by tuning paging thresholds, maintaining runbooks, and tracking postmortem actions.
  • Contribute to infrastructure build-outs and platform backlog work.
  • Participate in a shared on-call rotation of one week every six weeks.

Requirements

  • 5+ years of hands-on cloud or infrastructure engineering experience, including substantial production reliability work at scale.
  • Experience defining and operating SLIs and SLOs for production services.
  • Production experience with an observability platform; Datadog is strongly preferred.
  • Proficiency writing code in Python, Go, TypeScript, or a similar language for tooling, debugging, and automation.
  • Production Terraform experience and strong AWS expertise.
  • Experience creating and operating disaster recovery plans, including recovery objectives, failover procedures, restoration steps, and drills.
  • Experience being on call for services you helped build and developing effective alerting practices.
  • Ability to provide paved roads and sensible defaults that help engineering teams own their instrumentation.
  • Experience leading reliability or observability modernization projects (preferred).
  • Experience building internal tooling, running game days or chaos experiments, and managing observability costs (preferred).

Benefits

  • $150,000–$185,000 annual salary plus bonus and benefits.
  • Health, dental, and vision plans.
  • 401(k) matching.
  • Unlimited paid time off.
  • Daily lunch, snacks, coffee, and commuter benefits for in-office employees.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Field IT Operations Engineer

Green Thumb Industries 1K-5K Pharmaceuticals

Green Thumb is seeking a remote IT Operations Engineer to independently support end-user technology and core IT infrastructure across multiple business locations, with up to 40% travel and on-call participation.

iOS
4 hours, 31 minutes ago

Global IT Manager, Data Center IT

Serverfarm Data Centers 51-250 IT Services

Lead global site infrastructure operations at Serverfarm, a data center developer and operator, by managing distributed technical teams and ensuring reliable delivery, performance, and expansion of IT infrastructure across its international portfolio.

Linux Windows Server
4 hours, 46 minutes ago

IT Operations Coordinator 9/4/2026, 11:21:59 AM

Headway 251-1K Health Care Providers & Services

The Infrastructure Engineer will serve as a highly capable infrastructure specialist, supporting the organization’s remote operations.

1 day, 4 hours ago

Director, IT Engineering

The Director of IT Engineering at Ultimate Medical Academy provides strategic and technical leadership for secure, scalable enterprise engineering capabilities across cloud infrastructure, platform engineering, reliability, data, AI, and shared services.

AWS Azure CI/CD Cybersecurity Databricks DevSecOps HIPAA Terraform
1 day, 4 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers