Staff Data Engineer

4 months, 1 week ago
Full-time
Lead
Software Development
tvScientific

tvScientific

tvScientific offers a pioneering Connected TV advertising and marketing platform that integrates traditional television's impact with the effectiveness of digital advertising, specifically tailored for performance marketers.

Media
11-50
Founded 2020
$22M raised

Description

  • Design and maintain a scalable identity resolution platform.
  • Build pipelines and services to ingest, normalize, link, and version identity data from multiple sources.
  • Implement transparent, auditable, and measurable deterministic and probabilistic matching logic.
  • Partner with product and analytics teams to deliver reliable APIs and datasets for identity data.
  • Build and operate batch and streaming pipelines using modern data stack tools.
  • Create documentation, standards, and runbooks for identity and governance systems.
  • Own data governance foundations, including lineage, quality checks, schema enforcement, and access controls.
  • Implement privacy-by-design practices, including PII handling, consent enforcement, and retention policies.
  • Work with legal, privacy, and security teams to operationalize regulatory requirements such as GDPR and CCPA.
  • Establish monitoring and alerting for data quality, freshness, and integrity.

Requirements

  • Production data engineering experience.
  • Bachelor’s degree in computer science, a related field, or equivalent experience.
  • Proficiency in Spark and Scala, with experience building data infrastructure in Spark using Scala.
  • Experience delivering significant technical initiatives and building reliable, large-scale services.
  • Experience delivering APIs backed by relationship-heavy datasets.
  • Experience implementing data governance practices, including data quality, metadata management, and access controls.
  • Strong understanding of privacy-by-design principles and handling sensitive or regulated data.
  • Familiarity with data lakes, cloud warehouses, and storage formats.
  • Strong proficiency in AWS services.
  • Excellent written and verbal communication skills.
  • Demonstrated ability to design and implement scalable, efficient data infrastructure.
  • High attention to detail in implementing automated data quality checks.
  • Effective collaboration with cross-functional teams.
  • Demonstrated ability to use AI to improve speed and quality in day-to-day work for relevant outputs.
  • Strong track record of critically evaluating and verifying AI-assisted work through testing, source-checking, data validation, or peer review.
  • High integrity and ownership, with the ability to protect sensitive data and remain accountable for final decisions and deliverables.

Benefits

  • Base salary range of $177,185 to $364,795 USD for US-based applicants.
  • Eligible for equity.
  • Relocation assistance is not available for this position.
  • Work model information is available through PinFlex.
  • Benefits information for the role is available through Pinterest Life.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Lead Data Engineer

Media.Monks 5K-10K Media

Monks Technology Services is seeking a fully remote Lead Data Engineer for a fixed-term contract to lead the modernization of legacy Python/Spark ETL pipelines into dbt Core models on Databricks supporting analytics and business intelligence.

Apache Spark CI/CD Databricks Git GitHub Actions Python SQL
17 hours, 35 minutes ago

Staff Data Engineer

Webflow 251-1K Internet Software & Services

Webflow is hiring a Staff Data Engineer to lead data platform initiatives involving its data lake, event instrumentation, data quality, governance, and cloud infrastructure for modern web marketing operations.

Apache Airflow Apache Spark CI/CD Kafka SQL
17 hours, 35 minutes ago

Data Engineer [Zeal]

Livefront 11-50 Internet Software & Services

Zeal, now part of Livefront, is hiring a Data Engineer across its U.S. hubs to build reliable data pipelines and architectures that support technology consulting engagements for Fortune 1000 companies.

Apache Airflow CI/CD Databricks dbt GCP Git Kafka Oracle PostgreSQL Power BI Python RabbitMQ Snowflake SQL SQL Server Tableau
17 hours, 50 minutes ago

Sr Data Ops Engineer

Coderio 51-250 Internet Software & Services

Coderio is hiring a DataOps Engineer and Technical Referent to work with international customers on designing, implementing, and operating scalable cloud data infrastructure and automation solutions.

AWS Bash dbt Docker GitHub Python SQL Terraform
18 hours, 20 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers