PENN Entertainment

PENN Entertainment

PENN Entertainment is a leading provider of integrated entertainment, sports content, and casino gaming experiences across North America, offering a diverse range of entertainment destinations and a premier loyalty program that rewards members with cas...

Hotels, Restaurants & Leisure
10K-50K
Founded 1982

Description

  • Drive complex infrastructure migrations and projects across multiple production environments and jurisdictions.
  • Build and maintain platform tooling, GitOps workflows, release pipelines, and service onboarding automation.
  • Support development teams by consulting on infrastructure needs, reviewing architecture proposals, and resolving dependencies.
  • Design and improve Datadog observability, alerting, dashboards, and operational runbooks.
  • Investigate and resolve production incidents through structured debugging and root cause analysis.
  • Participate in on-call rotations and improve platform reliability.
  • Mentor teammates and reduce operational toil through automation and improved engineering practices.

Requirements

  • 5+ years of experience in Site Reliability Engineering, DevOps, or a similar role.
  • Strong production Kubernetes experience, including cluster lifecycle, networking, storage, and scheduling in Linux environments.
  • Experience with AWS, GCP, and/or on-premises infrastructure.
  • Proficiency in at least two of Go, Python, or Bash/Shell.
  • Deep understanding of distributed systems, networking, failure modes, capacity planning, and performance analysis.
  • Experience with GitOps and CI/CD tools such as ArgoCD, Helm, and GitHub Actions.
  • Experience with infrastructure as code using Terraform, Helm, or equivalent tools.
  • Track record leading complex infrastructure migrations or cross-team projects.
  • Strong incident response, troubleshooting, documentation, and technical communication skills.
  • Preferred: experience with service meshes, distributed storage, bare-metal Kubernetes or Talos OS, regulated environments, large-scale observability, or PostgreSQL operations.

Benefits

  • Salary range of $145,000–$193,000 USD.
  • Competitive compensation and comprehensive benefits package.
  • Remote work opportunity.
  • Education and conference reimbursements.
  • Bonus eligibility for most non-sales positions.
  • Support for employees’ physical, financial, and emotional well-being.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Staff Site Reliability Engineer, Ads

Reddit 1K-5K Internet Software & Services

Reddit is hiring a Staff Site Reliability Engineer to provide technical leadership for reliability, scalability, and operational excellence across its advertising infrastructure and revenue-critical systems.

Apache Spark ClickHouse GCP Go Kafka Kubernetes
16 hours, 43 minutes ago

Sr Lead Network Reliability Engineer

Coupa Software 1K-5K Internet Software & Services

Coupa is hiring a Sr. Lead Network Development Engineer to scale and operate its global SaaS platform’s cloud networking infrastructure through automation, reliability engineering, and technical leadership.

Ansible AWS Azure Chef DNS Fortinet Go Java Kubernetes Linux Python Ruby TCP/IP Terraform TLS
1 day, 16 hours ago

Senior Monitoring/Observability Architect

Makpar 51-250 Internet Software & Services

Makpar is seeking a Senior Monitoring/Observability Architect to lead enterprise monitoring strategy, architecture, and implementation guidance for a large federal government program.

Datadog Splunk
1 day, 16 hours ago

Staff Site Reliability Engineer

Caseware 251-1K Internet Software & Services

Caseware is hiring a Senior Site Reliability/Platform Engineer for its Canada-remote team to improve production resilience, security, operational excellence, and developer enablement across its audit and accounting software platform.

AWS AWS CDK CI/CD GitHub Actions Kubernetes Node.js OpenTelemetry TypeScript
1 day, 16 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers