Multi Media

Multi Media

Multi Media LLC is a leading live streaming technology company that provides innovative live streaming video solutions, engaging community forums, data insights, encryption, custom payment technology, global compliance, and smart in-app messages. Their...

Internet Software & Services
51-250
Founded 2011

Description

  • Analyze system performance using APM and distributed telemetry to identify instability sources.
  • Improve scalability, reliability, and performance through software enhancements and patching.
  • Develop tools and automation to streamline the DevOps pipeline.
  • Design and manage infrastructure across data center metal environments and public cloud platforms.
  • Conduct predictive failure analysis and disaster planning.
  • Administer and configure databases and key-value stores with a focus on uptime and performance.
  • Analyze complex systems to reduce operational surprises and minimize downtime.
  • Participate in incident response and write postmortem reports.
  • Collaborate with other engineering teams on reliability and infrastructure initiatives.
  • Help shape reliability strategy across systems and teams at different levels of ownership.

Requirements

  • STEM degree and/or relevant experience as a Site Reliability Engineer, DevOps Engineer, or Software Engineer.
  • Proficiency in Python or Golang, or another compiled/high-level language such as C, C#, C++, Java, or Rust.
  • Experience running web applications at scale.
  • Experience with web application concepts and frameworks such as ORM, MVC, Django, Flask, or Laravel.
  • Strong Linux administration skills, including Bash and knowledge of Linux internals such as filesystems and system calls.
  • Strong networking knowledge, including routing, switching, TCP stack, and cloud networking concepts such as VPCs and Security Groups.
  • Experience in database administration and configuration.
  • Experience with DevOps tools such as Terraform, Ansible, Docker, Kubernetes, ArgoCD, or Helm.
  • Willingness to participate in on-call rotation and respond to monitoring and alerting for core website functions.
  • Experience in production environments at an intermediate, senior, or staff level (preferred).

Benefits

  • Fair and competitive base salary of $169,000 - $215,000 USD.
  • Fully remote optional.
  • Health, vision, dental, and life insurance for you and dependents, with premiums covered by the company.
  • Long- and short-term disability insurance.
  • Unlimited PTO and 12 paid holidays.
  • Annual year-end company closure.
  • Optional 401(k) with 5% matching.
  • Paid lunches in-office or a $125/week remote stipend via Sharebite.
  • Employee Assistance and Employee Recognition Programs.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior SRE / Production Reliability Engineer

Symphony Solutions 251-1K Internet Software & Services

Senior SRE for a multi-brand iGaming platform at a company serving regulated markets worldwide, focused on improving the reliability, observability, and scalability of a microservices-based product stack.

Argo CD CI/CD Couchbase DNS Docker Elasticsearch GCP GitOps Grafana Helm Kafka Kubernetes Linux Load Balancing Microservices OpsGenie PagerDuty PostgreSQL Prometheus React Scala Terraform TLS TypeScript
18 minutes ago

Incident Commander

PENN Entertainment 10K-50K Hotels, Restaurants & Leisure

PENN Interactive is hiring an Incident Commander to join its site reliability team and lead incident response and service reliability efforts across its online and physical platforms.

Ansible AWS Docker Elasticsearch GCP Helm JIRA Kafka Kubernetes Linux MySQL PostgreSQL Prometheus Python Redis Terraform
1 day ago

Site Reliability Engineer

MyFitnessPal 10K-50K Health Care Providers & Services

MyFitnessPal is hiring a Software Engineer III, Site Reliability to own production reliability and delivery pipeline security for the PEAS team supporting automation, CI/CD, and self-service platforms.

AWS CI/CD Datadog GitHub Actions Go HIPAA Kubernetes Python Secrets Management Terraform TypeScript
1 day ago

Site Reliability Engineer 2 (Azure)

PhonePe 5K-10K Capital Markets

PhonePe Limited is hiring a Site Reliability Engineer to manage and scale core cloud infrastructure for a high-volume digital payments environment in India.

Ansible Azure Bash DNS Docker Go Grafana HAProxy InfluxDB Java Linux MySQL Nginx Prometheus Python RabbitMQ SaltStack Terraform Ubuntu
1 day, 23 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers