MyFitnessPal

MyFitnessPal

MyFitnessPal is a top nutrition tracking app with a calorie tracker, BMR calculator, and food tracker. It helps users reach health goals by tracking meals and physical activity, offering nonstop motivation for a healthier life.

Health Care Providers & Services
10K-50K
Founded 2005

Description

  • Own and evolve SLI/SLO and error-budget frameworks to inform prioritization and product decisions.
  • Lead incident response, drive postmortems, and implement systemic fixes.
  • Build and maintain observability across metrics, logs, and traces using Datadog.
  • Design and operate resilient, scalable infrastructure using Terraform.
  • Manage production Kubernetes and container workloads, including capacity planning and cloud-cost optimization.
  • Own CI/CD pipelines and safe deployment strategies such as canary, progressive rollout, and fast rollback.
  • Integrate and tune security scanning in the delivery pipeline, including SAST, DAST, and SCA.
  • Implement and maintain policy-as-code controls to block unsafe infrastructure and Kubernetes changes at admission time.
  • Drive vulnerability triage and remediation SLAs for pipeline and infrastructure findings.
  • Improve on-call operations by building sustainable runbooks and automation, and coach engineers on reliability best practices.

Requirements

  • 5+ years of experience in site reliability, platform, or infrastructure engineering with senior-level ownership of production systems.
  • Strong programming skills for automation and tooling in Go, Python, TypeScript, or similar languages.
  • Deep hands-on experience with a major cloud platform, Kubernetes, and Infrastructure as Code.
  • Experience with AWS is a plus.
  • Experience with Terraform is a plus.
  • Proven track record leading incident response and building SLO-driven reliability practices.
  • Working fluency with observability tooling, with Datadog as a plus.
  • Practical experience integrating security into CI/CD pipelines, including SAST/DAST/SCA, dependency scanning, or policy-as-code.
  • Strong understanding of cloud security fundamentals, including IAM, least privilege, policy guardrails, and secrets management.
  • Experience with Kyverno, OPA/Rego, or Conftest enforced at admission time is a plus.
  • Exposure to regulated or compliance-driven environments such as SOC 2, PCI DSS, or HIPAA is a plus.
  • Chaos engineering or game-day experience is a plus.
  • Experience supporting B2C or mobile backend environments with high traffic and strong reliability needs is a plus.

Benefits

  • Salary range of $120,000-$165,000.
  • Annual performance bonus.
  • Comprehensive healthcare benefits, including medical, dental, and vision.
  • Parental planning support, including paid maternity and paternity leave and fertility assistance.
  • 401(k) retirement plan with employer match.
  • Responsible time off policy.
  • Monthly wellness and technology allowances.
  • Mental health benefits and dedicated mental health days.
  • Access to MyFitnessPal Premium.
  • Learning and development resources and training opportunities.
  • Volunteer days off.
  • Flexible, in-person team connection opportunities and annual company gatherings.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Site Reliability Engineer

AssureSoft 51-250 Internet Software & Services

AssureSoft is hiring a remote Site Reliability Engineer to support production cloud infrastructure and platform reliability for long-term client projects.

Argo CD AWS Bash DNS Docker GCP GitHub Actions Go Grafana HTTP Kafka Kubernetes Linux Load Balancing Prometheus Python RabbitMQ Snowflake TCP/IP TLS TypeScript Unix
1 day ago

Senior Site Reliability Engineer

Megaport 251-1K Diversified Telecommunication Services

Megaport is hiring a Senior Platform Engineer to support secure, reliable, and maintainable global production systems within its SRE-focused platform team.

AWS Bash Cassandra CI/CD ClickHouse Git GitHub Go Kubernetes Linux PostgreSQL Python Terraform
1 day, 1 hour ago

Site Reliability Engineer - Azure, Observability and Scripting

Jalasoft 1K-5K Internet Software & Services

Jalasoft is hiring a Site Reliability Engineer to support the reliability, scalability, and performance of cloud-native platforms on Microsoft Azure and Kubernetes.

Azure Bash Grafana Kubernetes OpenTelemetry Oracle PowerShell Prometheus Python Terraform
2 days, 1 hour ago

Senior Service Reliability Engineer

Thoughtworks 10K-50K Professional Services

Senior Service Reliability Engineer at Thoughtworks, focused on improving infrastructure reliability, observability, and incident response for customer-facing production systems.

Azure Bash Datadog ELK Stack GitOps Go Grafana Kubernetes Microservices Network Security New Relic Nomad Python REST API Serverless Terraform
2 days, 1 hour ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers