GiveCampus

GiveCampus

GiveCampus is the world's #1 fundraising platform for nonprofit educational institutions, offering a blend of crowdfunding, peer-to-peer fundraising, social networking, and gamification. Trusted by over 1,000 colleges, universities, and K-12 schools, G...

Internet Software & Services
51-250
Founded 2014

Description

  • Operate, maintain, and improve production infrastructure in AWS, including Kubernetes and Amazon EKS workloads.
  • Build and maintain infrastructure as code with Terraform.
  • Improve dashboards, alerts, and service-level indicators using New Relic or comparable observability tools.
  • Investigate production incidents, identify root causes, implement durable fixes, and participate in the 24/7 on-call rotation.
  • Lead blameless postmortems and complete follow-up actions to prevent recurring incidents.
  • Partner with product engineers to troubleshoot application performance and reliability issues.
  • Improve application resilience through timeouts, retries, queuing, backpressure, and idempotency.
  • Maintain CI/CD pipelines and deployment workflows using tools such as GitHub Actions and CircleCI.
  • Automate operational tasks, reduce engineering toil, and maintain runbooks, diagrams, troubleshooting guides, and production documentation.
  • Contribute to capacity planning, performance testing, database reliability, security practices, and production-readiness reviews.

Requirements

  • Approximately 5+ years of experience in SRE, infrastructure, systems engineering, software engineering, platform engineering, DevOps, or equivalent practical experience.
  • Hands-on experience operating production workloads in AWS.
  • Experience with Terraform or a similar infrastructure-as-code tool.
  • Experience with New Relic, Datadog, or another modern observability platform.
  • Experience troubleshooting production incidents and participating in an on-call rotation.
  • Experience building or maintaining CI/CD pipelines.
  • Software development or scripting experience, including reading, debugging, and modifying application or automation code.
  • Working knowledge of Linux, networking, distributed systems, and relational databases.
  • Strong written and verbal communication skills, including explaining root causes and technical tradeoffs.
  • Experience with Ruby or Ruby on Rails, PostgreSQL administration, Kubernetes/EKS, Redis, OpenSearch, or Amazon RDS is preferred.
  • Familiarity with SLOs, SLIs, error budgets, capacity modeling, load testing, regulated systems, or SOC 2 programs is preferred.

Benefits

  • Remote-first work environment for U.S.-based employees.
  • Flexible access to a 12,000-square-foot Washington, DC office.
  • Company-wide and team-specific onsite events, meetups, partner visits, and retreats.
  • Mission-driven work supporting educational institutions and access to education.
  • Inclusive, supportive environment focused on learning and professional growth.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Cloud Site Reliability Engineer

Cadwell 51-250 Health Care Providers & Services

Cadwell is seeking a Cloud Site Reliability Engineer to operate and improve AWS infrastructure supporting healthcare customers and ensure reliable, secure, and compliant hosted neurodiagnostic software environments.

AWS Bash CI/CD Encryption HIPAA JavaScript JSON Python Terraform TypeScript YAML
25 minutes ago

DevOps / SRE / DevSecOps Engineer (AWS) - Latin America, Remote

Bluelight Consulting 11-50 Internet Software & Services

Bluelight is hiring a DevOps/SRE/DevSecOps professional to operate secure, reliable AWS-based systems and delivery infrastructure for client software projects in a remote consultancy environment.

AWS CI/CD DevSecOps Docker K6 OpenTelemetry PostgreSQL Secrets Management Terraform
1 day ago

DevOps / SRE / DevSecOps Engineer (AWS) - Latin America, Remote

Bluelight Consulting 11-50 Internet Software & Services

Bluelight is hiring a DevOps/SRE/DevSecOps professional to support complex client systems by building secure, reliable, and observable AWS infrastructure and delivery operations.

AWS CI/CD DevSecOps Docker K6 OpenTelemetry PostgreSQL Secrets Management Terraform
1 day ago

DevOps / SRE / DevSecOps Engineer (AWS) - Latin America, Remote

Bluelight Consulting 11-50 Internet Software & Services

Bluelight is hiring a DevOps/SRE professional to support complex, security-sensitive client systems by building reliable AWS infrastructure, deployment processes, and operational practices within its remote software consultancy.

AWS CI/CD DevSecOps Docker K6 OpenTelemetry PostgreSQL Secrets Management Terraform
1 day ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers