Simulmedia

Simulmedia

Simulmedia is a pioneer in cross-channel TV advertising, offering advertisers the reach of linear TV combined with digital targeting. Their VAMOS platform enables intelligent audience optimization and direct sales attribution, revolutionizing TV advert...

Media
51-250
Founded 2008
$87M raised

Description

  • Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets from external providers and internal systems.
  • Model complex data structures, including dimensional, reference, and temporal models, and evolve them safely as upstream schemas change.
  • Develop and operate workloads on Databricks/Spark/Delta and Redshift, including migrating pipelines to the lakehouse platform.
  • Orchestrate pipelines in Airflow, including scheduling, dependencies, retries, backfills, and alerting.
  • Design parity checks and reconciliation queries, run historical backfills, and investigate data discrepancies at the row level.
  • Build and maintain Python services and REST APIs that serve data to internal products.
  • Optimize performance and cost through query tuning, table design, workload management, and compute right-sizing.
  • Monitor production pipelines, participate in incident triage and root-cause analysis, and harden systems against repeat failures.
  • Collaborate with product managers, data scientists, and other stakeholders to deliver roadmap items.
  • Work in an Agile team and experiment with new technologies to improve software and data workflows.

Requirements

  • Bachelor's degree in Computer Science, Computer Engineering, a relevant technical field, or equivalent practical experience.
  • 7+ years of work experience as a data engineer.
  • Proficiency in Python as the primary development language in recent years.
  • Expert-level SQL for complex analytical queries, performance tuning, and debugging result discrepancies.
  • Hands-on experience with a distributed data processing platform such as Spark/Databricks, EMR, Snowflake, or BigQuery.
  • Experience with a columnar data warehouse such as Redshift, Snowflake, BigQuery, or ClickHouse.
  • Ability to design complex data models, including normalized, dimensional, and temporal models.
  • Experience with workflow orchestration tools such as Airflow, including DAGs, dependencies, and backfills.
  • Experience integrating third-party data feeds, including schema drift, late deliveries, missing data, and vendor data-quality defects.
  • Experience building REST services in Python using FastAPI, Flask, or similar frameworks.
  • Working knowledge of AWS services such as S3, IAM, and ECS, plus Docker.
  • Good knowledge of engineering best practices, including unit testing, integration testing, code review, and CI/CD.
  • Must be able to communicate with U.S.-based teams and work 11:00 AM — 8:00 PM EEST.
  • Experience with Delta Lake or medallion lakehouse architectures is a plus.
  • Experience migrating legacy pipelines between platforms with strict parity requirements is a plus.
  • Experience with advertising, media, or measurement industry data is a plus.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer

UJET 251-1K Professional Services

UJET is hiring a Data Engineer to own and modernize the end-to-end data platform powering customer analytics, operational reporting, and trusted business metrics.

CI/CD dbt GCP Git Go Kubernetes Looker MySQL Python Ruby on Rails SQL Terraform
3 hours, 32 minutes ago

Data and Machine Learning Intern

Loka 51-250 Internet Software & Services

Loka is hiring a full-time, paid six-month Data–Machine Learning Intern to support global teams delivering data and machine learning projects while building practical experience and a project portfolio.

Apache Airflow Apache Spark AWS Deep Learning Machine Learning Python
4 hours, 17 minutes ago

Data Engineer

EP Wealth 251-1K Capital Markets

EP Wealth Advisors is seeking a hands-on Data Engineer to build and operate a scalable Snowflake- and Azure-based enterprise data platform supporting analytics, reporting, operational workflows, and future AI capabilities.

Azure CI/CD Git Machine Learning Python REST API Salesforce Secrets Management Snowflake SQL
1 day, 3 hours ago

[Job-31836] Senior Data Developer, Brazil

CI&T 5K-10K Internet Software & Services

CI&T is hiring Senior Data Engineers to modernize a Finance data platform for a major US mortgage lender by converting raw replicated data into governed, reconciled, business-defined Finance measures.

CI/CD Git Power BI Snowflake SQL
1 day, 4 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers