Sezzle

Sezzle

Sezzle is a payments company revolutionizing the buy now, pay later experience with interest-free installment plans, empowering consumers and merchants alike.

Diversified Financial Services
251-1K
Founded 2016

Description

  • Own the architecture and evolution of core infrastructure, identifying limits and prioritizing improvements.
  • Design and build resilient AWS account, IAM, networking, and multi-region or multi-AZ architectures.
  • Operate and improve Kubernetes platforms, including lifecycle automation, workload isolation, autoscaling, upgrades, and deployment reliability.
  • Scale and optimize Aurora RDS for MySQL and Postgres through query tuning, capacity planning, replication improvements, and safe migrations.
  • Engineer reliability through service-level objectives, failure isolation, backpressure, load shedding, and safe retry strategies.
  • Participate in on-call rotations, lead recovery during major incidents, and implement postmortem corrective actions.
  • Build disaster recovery, backup, restore, failover, observability, infrastructure-as-code, and operational automation capabilities.
  • Conduct capacity planning, load testing, bottleneck analysis, safe migrations, and cloud cost optimization.
  • Develop and evaluate AI-assisted tooling for incident investigation, runbooks, anomaly analysis, and toil reduction.
  • Write architecture proposals, prototype and benchmark solutions, review shared infrastructure changes, and document system behavior.

Requirements

  • Bachelor’s degree in Computer Science or a similar technical field.
  • 12+ years of experience across infrastructure, platform, SRE, software development, or related engineering disciplines.
  • Deep production expertise with AWS, including compute, IAM, multi-account architecture, VPC design, and private networking.
  • Deep production expertise with Kubernetes; EKS experience is strongly preferred.
  • Experience with RDS/Aurora MySQL and/or Postgres at scale, including performance tuning, replication, high availability, failover, and recovery.
  • Track record of personally delivering measurable infrastructure improvements in capacity, latency, reliability, or cost.
  • Strong coding and automation skills in Golang, Python, or similar languages, plus Terraform or equivalent infrastructure-as-code experience.
  • Strong fundamentals in Linux, networking, DNS, TLS, storage, concurrency, and distributed-system failure modes.
  • Experience operating 24/7 high-availability platforms, participating in incident response, and testing disaster recovery.
  • Experience with observability, load testing, capacity planning, safe CI/CD, and active use of AI tools with appropriate validation.
  • Preferred: fintech, payments, or banking experience; multi-region architectures; chaos engineering; Prometheus, Grafana, Loki, or Tempo; and internal platform tooling.

Benefits

  • Monthly gross compensation of $12,500–$20,800 USD, based on location and experience.
  • Remote full-time role.
  • Opportunity to work on fintech infrastructure with AWS, Kubernetes, Aurora, Golang, and Python.
  • Opportunity to build AI-assisted infrastructure and SRE tooling.
  • Collaborative environment focused on open-source technology and measurable engineering impact.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Staff Cloud Architect L4

Robots & Pencils 51-250 IT Services

Robots & Pencils is seeking a Staff Cloud Architect for a six-month contract to lead end-to-end architecture and delivery of complex cloud, data, and AI/ML systems for enterprise client engagements.

Apache Airflow Apache Spark AWS Azure CI/CD dbt Docker GitHub Actions HIPAA Hugging Face Kafka Kubernetes Machine Learning Microservices MLflow MLOps MongoDB PostgreSQL Prefect Pulumi Python PyTorch Serverless Snowflake TensorFlow Terraform Vertex AI
56 minutes ago

Infrastructure Software Engineer, Metadata Services

Dropbox 1K-5K Internet Software & Services

Software Engineer on Dropbox’s Metadata team, building and operating distributed database infrastructure that provides reliable, low-latency, strongly consistent data access for millions of users and internal services.

C C++ Go Java Python Rust
1 hour, 26 minutes ago

Senior Infrastructure Engineer

Sysdig 251-1K IT Services

As a Site Reliability/Infrastructure Engineer at Sysdig, you will build and operate multi-cloud and on-premise infrastructure while improving the reliability, scalability, security, and performance of production systems.

AWS Azure Bash Docker Go Kubernetes Linux Microservices Python
1 hour, 41 minutes ago

Cloud Security Engineer

Smile Digital Health 251-1K IT Services

Smile Digital Health is seeking a Cloud Security Engineer to design, automate, deploy, and support secure production-grade infrastructure for healthcare data platforms across AWS, Azure, OCI, and GCP.

Ansible AWS Azure Docker HIPAA Kubernetes OpenShift Terraform
2 days ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers