MRSOOL

MRSOOL

Mrsool: Leading on-demand delivery platform with a vast courier fleet, high user ratings, and innovative chat-based ordering. Established in 2015, serving 10 million users in Saudi Arabia.

Air Freight & Logistics
1K-5K
Founded 2015

Description

  • Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt.
  • Develop and optimize data models based on Medallion Architecture to deliver reliable, reusable, high-quality datasets.
  • Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery.
  • Design data ingestion frameworks using CDC, Kafka, and event-driven architectures for near real-time processing.
  • Create and maintain data warehouses and data marts for reporting and self-service analytics.
  • Partner with Product Managers, Data Analysts, Backend Engineers, and Business stakeholders to translate requirements into data solutions.
  • Develop reusable dbt models, automated tests, and documentation to improve data quality and developer productivity.
  • Optimize Spark jobs, Trino queries, and storage layouts for performance, reliability, and cost efficiency.
  • Own the end-to-end lifecycle of critical data pipelines, including monitoring, SLA adherence, and incident resolution.
  • Build automation, CI/CD pipelines, and reusable frameworks to strengthen the core data platform.

Requirements

  • 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses.
  • Strong proficiency in Spark (Scala, Python) and SQL.
  • Hands-on experience with Apache Spark and distributed data processing, performance tuning, and optimization.
  • Experience building batch and streaming data pipelines using Kafka, CDC/Maxwell, or similar event-driven architectures.
  • Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization.
  • Experience with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines.
  • Solid experience designing dimensional models, star schemas, and reliable data marts.
  • Hands-on experience with dbt, including reusable models, automated testing, and documentation.
  • Strong knowledge of data quality, observability, lineage, and engineering best practices.
  • Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs for performance, scalability, and cost efficiency.
  • Familiarity with CI/CD, Git-based workflows, infrastructure automation, and modern software engineering best practices.
  • Excellent problem-solving skills with the ability to independently own projects from design through production.
  • Strong communication and stakeholder management skills across Product, Engineering, Analytics, and Business teams.

Benefits

  • Inclusive and diverse workplace with remote work environments.
  • Competitive compensation packages.
  • Potential share options for certain roles.
  • Regular training and an annual learning stipend.
  • High degree of autonomy in the role.
  • Mentorship and ambitious goals that support growth.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Lead Data Engineer

Media.Monks 5K-10K Media

Monks Technology Services is seeking a fully remote Lead Data Engineer for a fixed-term contract to lead the modernization of legacy Python/Spark ETL pipelines into dbt Core models on Databricks supporting analytics and business intelligence.

Apache Spark CI/CD Databricks Git GitHub Actions Python SQL
17 hours, 28 minutes ago

Staff Data Engineer

Webflow 251-1K Internet Software & Services

Webflow is hiring a Staff Data Engineer to lead data platform initiatives involving its data lake, event instrumentation, data quality, governance, and cloud infrastructure for modern web marketing operations.

Apache Airflow Apache Spark CI/CD Kafka SQL
17 hours, 28 minutes ago

Data Engineer [Zeal]

Livefront 11-50 Internet Software & Services

Zeal, now part of Livefront, is hiring a Data Engineer across its U.S. hubs to build reliable data pipelines and architectures that support technology consulting engagements for Fortune 1000 companies.

Apache Airflow CI/CD Databricks dbt GCP Git Kafka Oracle PostgreSQL Power BI Python RabbitMQ Snowflake SQL SQL Server Tableau
17 hours, 43 minutes ago

Sr Data Ops Engineer

Coderio 51-250 Internet Software & Services

Coderio is hiring a DataOps Engineer and Technical Referent to work with international customers on designing, implementing, and operating scalable cloud data infrastructure and automation solutions.

AWS Bash dbt Docker GitHub Python SQL Terraform
18 hours, 13 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers