MRSOOL

MRSOOL

Mrsool: Leading on-demand delivery platform with a vast courier fleet, high user ratings, and innovative chat-based ordering. Established in 2015, serving 10 million users in Saudi Arabia.

Air Freight & Logistics
1K-5K
Founded 2015

Description

  • Collaborate with development teams to design and implement scalable infrastructure.
  • Collaborate with development teams to design and implement automated deployment and testing pipelines.
  • Develop and maintain monitoring and alerting systems to identify and address issues proactively.
  • Troubleshoot and escalate production incidents to minimize downtime and improve system reliability.
  • Continuously improve infrastructure and processes to optimize scalability and efficiency.
  • Participate in on-call rotations to provide 24/7 application support.
  • Perform routine maintenance and upgrades to keep systems current.
  • Contribute to improvements in security posture and compliance with industry standards.
  • Communicate complex technical concepts clearly to technical and non-technical stakeholders.
  • Mentor and coach junior engineers to support their growth and delivery quality.

Requirements

  • Bachelor’s degree in Computer Engineering, Computer Science, or a related field.
  • 5+ years of experience in a similar role, preferably in a high-traffic, high-availability environment.
  • Proficiency in at least one programming language such as Python, Ruby, Java, or Go.
  • Strong understanding of cloud infrastructure and technologies such as AWS, GCP, Azure, Kubernetes, and Docker.
  • Experience with automation and configuration management tools such as Chef, Ansible, Puppet, or Terraform.
  • Familiarity with monitoring and alerting tools such as Prometheus, Grafana, or Nagios.
  • Excellent troubleshooting and problem-solving skills.
  • Strong communication and interpersonal skills for collaboration with cross-functional teams.
  • Ability to navigate ambiguity, set clear expectations, and work in a fast-paced, dynamic environment.
  • Strong grasp of computer science fundamentals related to distributed systems and networks.
  • Experience running, tuning, and optimizing databases and queries is preferred.
  • Prior leadership or senior-level experience in site reliability engineering is preferred.
  • Familiarity with additional backend technologies and frameworks is preferred.
  • Experience driving technical decisions and organizational improvements is preferred.

Benefits

  • Inclusive and diverse work environment.
  • Remote work environment.
  • Competitive compensation package.
  • Potential share options for certain roles.
  • Regular training and an annual learning stipend.
  • High degree of autonomy in the role.
  • Mentorship and ambitious goals to support growth.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its document workflow platform reliable, resilient, and available for customers.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
16 hours, 23 minutes ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its document workflow platform reliable through incident management, observability, production support, and resilience work across services.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
16 hours, 23 minutes ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to keep its document workflow platform highly available and resilient while supporting production operations and reliability improvements.

Agile AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
1 day, 15 hours ago

Senior Site Reliability Engineer

PandaDoc 251-1K Internet Software & Services

PandaDoc is hiring a Site Reliability Engineer to help keep its production document workflow platform reliable, resilient, and low-downtime for customers.

AWS Django Grafana Java Kafka Kubernetes NATS PostgreSQL Python RabbitMQ Spring Boot
1 day, 16 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers