Simulmedia

Simulmedia

Simulmedia is a pioneer in cross-channel TV advertising, offering advertisers the reach of linear TV combined with digital targeting. Their VAMOS platform enables intelligent audience optimization and direct sales attribution, revolutionizing TV advert...

Media
51-250
Founded 2008
$87M raised

Description

  • Design and build batch data pipelines that ingest, validate, and transform multi-billion-row datasets from external providers and internal systems.
  • Model complex data structures, including dimensional, reference, and temporal models, and evolve them safely as upstream schemas change.
  • Develop and operate workloads on Databricks/Spark/Delta and Redshift, including migrating pipelines to the lakehouse platform.
  • Orchestrate pipelines in Airflow, including scheduling, dependencies, retries, backfills, and alerting.
  • Design parity checks and reconciliation queries, run historical backfills, and investigate data discrepancies at the row level.
  • Build and maintain Python services and REST APIs that serve data to internal products.
  • Optimize performance and cost through query tuning, table design, workload management, and compute right-sizing.
  • Monitor production pipelines, participate in incident triage and root-cause analysis, and harden systems against repeat failures.
  • Collaborate with product managers, data scientists, and other stakeholders to deliver roadmap items.
  • Work in an Agile team and experiment with new technologies to improve software and data workflows.

Requirements

  • Bachelor's degree in Computer Science, Computer Engineering, a relevant technical field, or equivalent practical experience.
  • 7+ years of work experience as a data engineer.
  • Proficiency in Python as the primary development language in recent years.
  • Expert-level SQL for complex analytical queries, performance tuning, and debugging result discrepancies.
  • Hands-on experience with a distributed data processing platform such as Spark/Databricks, EMR, Snowflake, or BigQuery.
  • Experience with a columnar data warehouse such as Redshift, Snowflake, BigQuery, or ClickHouse.
  • Ability to design complex data models, including normalized, dimensional, and temporal models.
  • Experience with workflow orchestration tools such as Airflow, including DAGs, dependencies, and backfills.
  • Experience integrating third-party data feeds, including schema drift, late deliveries, missing data, and vendor data-quality defects.
  • Experience building REST services in Python using FastAPI, Flask, or similar frameworks.
  • Working knowledge of AWS services such as S3, IAM, and ECS, plus Docker.
  • Good knowledge of engineering best practices, including unit testing, integration testing, code review, and CI/CD.
  • Must be able to communicate with U.S.-based teams and work 11:00 AM — 8:00 PM EEST.
  • Experience with Delta Lake or medallion lakehouse architectures is a plus.
  • Experience migrating legacy pipelines between platforms with strict parity requirements is a plus.
  • Experience with advertising, media, or measurement industry data is a plus.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer

Effectual 51-250 Internet Software & Services

Effectual is hiring a Data Engineer to design, build, and maintain data architecture and pipelines that support business analytics, reporting, and data science work.

Agile Apache Airflow Apache Spark AWS Azure Computer Vision Databricks GCP Hadoop Java Python Scala Snowflake SQL Terraform
9 hours, 15 minutes ago

Azure Data Platform Engineer

Xebia 1K-5K Internet Software & Services

Xebia is seeking a platform engineer to help design and migrate modern data platform environments on Azure, Databricks, and Microsoft Fabric for enterprise clients.

Apache Airflow Azure CI/CD Databricks dbt Kubernetes Python Terraform
9 hours, 45 minutes ago

Analytics Platform Engineer

Coforge 10K-50K IT Services

Coforge is hiring a remote Analytics Platform Engineer in Brazil to support and improve the reliability, operations, and performance of its analytics data platform.

Apache Airflow AWS GitHub Actions Snowflake SQL
9 hours, 45 minutes ago

Snowflake Data Engineer

FreedomPay 251-1K IT Services

FreedomPay is seeking a Snowflake Data & AI Engineer to deliver production-grade data, analytics, and AI solutions for its global commerce platform.

Azure CI/CD JavaScript Kafka LLM Machine Learning NetSuite Python Snowflake SQL SQL Server
10 hours, 45 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers