Manager, Software Engineering (Reliability Platform)

3 weeks, 3 days ago
Full-time
Lead
Software Development
Affirm

Affirm

Affirm offers a transparent buy now, pay later service founded in 2012 by Max Levchin. No late fees or surprises, just a responsible way to pay over time for your favorite brands.

Diversified Financial Services
1K-5K
Founded 2012

Description

  • Build and lead a high-performing product engineering team with a culture of innovation, accountability, reliability, and execution.
  • Leverage AI-assisted engineering workflows to accelerate platform development, operational automation, and engineering productivity.
  • Translate ambiguous operational and organizational challenges into clear technical requirements, execution plans, and measurable outcomes.
  • Develop scalable reliability, risk management, and operational governance capabilities for Affirm’s production systems using observability tooling, automation, and AI.
  • Build foundational systems that help engineers understand service ownership, dependencies, and operational topology across Affirm’s infrastructure and applications.
  • Drive alignment across Platform Engineering, SRE, Infrastructure, and product engineering teams on requirements, investment priorities, and long-term roadmap outcomes.
  • Lead delivery of critical cross-functional reliability initiatives with high organizational visibility.
  • Partner closely with Infrastructure and Product Engineering teams to define reliability standards and build scalable operational tooling.

Requirements

  • 7+ years of experience in backend or full stack engineering.
  • 2+ years of engineering leadership experience.
  • Experience working within or alongside Production Engineering or Site Reliability Engineering teams.
  • Expertise with observability tools and with building software or managing programs that drive engineering culture and reliability practices.
  • Strong operational judgment and ability to drive clarity, prioritization, and execution in ambiguous problem spaces.
  • Experience building internal platforms, developer tooling, reliability products, or operational automation systems at scale.
  • Demonstrated ability to balance rapid iteration and experimentation with operational rigor and long-term maintainability.
  • Strong programming background in Python, Kotlin, Java, or similar languages.
  • Strong communication and organizational leadership skills, including aligning stakeholders, driving execution, and influencing engineering practices across teams.
  • Equivalent practical experience or a Bachelor’s degree in a related field is required.

Benefits

  • Remote-first work environment with flexibility to work almost anywhere within the country of employment.
  • Competitive base salary of $230,000 - $290,000 in CA, WA, NY, NJ, and CT, or $204,000 - $264,000 in other U.S. states.
  • Equity rewards, including Equity Grade 13.
  • Monthly stipends for health, wellness, and tech spending.
  • 100% subsidized medical, dental, and vision coverage for you and your dependents.
  • Flexible Spending Wallets for technology, food, lifestyle needs, and family-forming expenses.
  • Competitive vacation and holiday schedules.
  • Employee stock purchase plan with discounted shares.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Site Reliability Engineer III

onXmaps 251-1K Food Products

onX is hiring a Site Reliability Engineer to manage the infrastructure, deployment automation, and observability that help developers ship reliably at scale for its outdoor technology products.

Apache Airflow CockroachDB GCP Kubernetes OpenTelemetry Prometheus SQL Terraform
13 hours, 49 minutes ago

Site Reliability Engineer - South Korea

MinIO 51-250 Internet Software & Services

MinIO is hiring a Site Reliability Engineer to help enhance and operate its cloud-native storage platform for high-performance, scalable, and durable data storage and retrieval.

C C++ GitOps Go Kubernetes Microservices Rust
13 hours, 49 minutes ago

Senior Site Reliability Engineer- FedRamp

Veeam Software 1K-5K Internet Software & Services

Veeam is hiring a Site Reliability Engineer to help build its global SRE function for the Veeam Data Cloud, focused on the Government and Sovereign Cloud environment.

Argo CD Azure Bitbucket C# CI/CD ELK Stack Git GitHub Actions GitLab CI GitOps Go Grafana HIPAA Java JavaScript Kubernetes OpenTelemetry Prometheus Pulumi Terraform TypeScript
14 hours, 49 minutes ago

Senior Engineering Manager, Servicing

Upstart 1K-5K Banks

Upstart is hiring a Senior Engineering Manager to lead the GBOT team in building scalable, compliant systems that help borrowers get back on track while partnering across Product, Risk, Operations, and Machine Learning.

Machine Learning
15 hours, 4 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers