Ripjar

Ripjar

Ripjar provides advanced data analytics solutions for global companies and governments to detect criminal behavior, measure campaign success, and ensure AML compliance in real time.

IT Services
51-250
Founded 2012
$42M raised

Description

  • Engineer distributed ingestion services that pull data from diverse sources and produce clean, structured outputs for downstream systems.
  • Build high-throughput batch and near-real-time processing components with a focus on performance, scalability, and predictable cost.
  • Design and evolve data contracts, including schemas, validation rules, versioning, and backward compatibility.
  • Own production quality through maintainable code, unit and integration tests, and observability using metrics, logs, and tracing.
  • Harden pipelines against partial failures, retries, rate limits, data drift, and infrastructure issues.
  • Improve CI/CD and developer experience through faster builds, better test signals, safer releases, and automated operational checks.
  • Participate in design reviews, code reviews, incident retrospectives, and iterative delivery.
  • Document pragmatic trade-offs and implementation decisions clearly.
  • Work across distributed systems and operational reliability in collaboration with downstream product teams.

Requirements

  • 2+ years of experience building and operating production software systems.
  • Fluency in at least one programming language; Python or Node.js is a plus.
  • Experience debugging moderately complex systems and improving reliability and performance.
  • Strong fundamentals in data structures, testing, version control, and Linux basics.
  • Experience with Spark or PySpark is preferred.
  • Exposure to the Hadoop ecosystem, including HDFS and HBase, is preferred.
  • Experience with workflow orchestration tools such as Airflow, Dagster, or NiFi is preferred.
  • Experience with search/indexing systems such as OpenSearch or MongoDB is preferred.
  • Experience with Kubernetes and infrastructure-as-code is preferred.
  • A degree in Computer Science or a numerical degree is preferred.

Benefits

  • Competitive salary, dependent on experience.
  • 25 days annual leave plus your birthday off, in addition to bank holidays, rising to 30 days after 5 years of service.
  • Remote working.
  • Private family healthcare.
  • 35-hour working week.
  • Employee Assistance Programme.
  • Company pension contributions and pension salary sacrifice.
  • Enhanced maternity and paternity pay.
  • Latest tech, including a top-of-the-range MacBook Pro.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Data Engineer (Contract)

Concurrency 51-250 Internet Software & Services

Concurrency is hiring a Data Engineer for its Cloud, Data & AI Solutions Team to deliver end-to-end Microsoft Azure and Fabric data solutions that support client digital transformation.

Apache Spark Azure CI/CD Git Machine Learning Power BI Python SQL
42 minutes ago

Data Engineer II

Netomi 51-250 IT Services

Netomi is hiring a Data Engineer in Gurugram to build reliable data infrastructure and analytics solutions that improve AI-powered enterprise customer experiences and product operations.

Apache Airflow Apache Spark CI/CD Databricks Docker HIPAA Kafka Kubernetes Luigi MySQL PostgreSQL Prefect Python RabbitMQ Snowflake SQL
42 minutes ago

[Job-31682] Senior Data Engineer [Databricks]

CI&T 5K-10K Internet Software & Services

A CI&T busca uma pessoa engenheira de dados para estruturar e acelerar a plataforma de dados de um grande varejista farmacêutico brasileiro, modernizando a orquestração, a qualidade dos dados e a distribuição de produtos baseados em Databricks.

Apache Airflow Apache Spark AWS Databricks SFTP Terraform
57 minutes ago

Data Engineer

payabl. 51-250 Diversified Financial Services

As a Data Engineer at payabl., you will build and maintain reliable data pipelines and business-ready datasets that support analytics and decision-making across the company’s global payments platform.

Apache Airflow Apache Spark AWS ClickHouse Dagster Databricks dbt Docker Git Kafka Kubernetes MariaDB MongoDB MySQL PostgreSQL Power BI Python Snowflake SQL Tableau Terraform
57 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers