Software Engineer II, Big Data

2 months, 2 weeks ago
Full-time
Mid Level
Software Development
tvScientific

tvScientific

tvScientific offers a pioneering Connected TV advertising and marketing platform that integrates traditional television's impact with the effectiveness of digital advertising, specifically tailored for performance marketers.

Media
11-50
Founded 2020
$22M raised

Description

  • Design and implement robust data infrastructure in AWS using Spark and Scala.
  • Evolve core data pipelines to scale efficiently with company growth.
  • Store data in optimal engines and formats to balance performance and cost.
  • Collaborate with cross-functional teams to design data solutions that meet business needs.
  • Design and implement knowledge graphs and expose them through batch processing and APIs.
  • Leverage and optimize AWS resources while designing for scale.
  • Work closely with Data Science and Product teams.
  • Implement automated data quality checks with strong attention to detail.
  • Define and implement a strategic vision for data engineering as an individual contributor.

Requirements

  • Production data engineering experience.
  • Proficiency in Spark and Scala, with preferred experience building data infrastructure in Spark using Scala.
  • Experience delivering significant technical initiatives and building reliable, large-scale services.
  • Experience delivering APIs backed by relationship-heavy datasets.
  • Familiarity with data lakes, cloud warehouses, and storage formats.
  • Strong proficiency in AWS services.
  • Expertise in SQL for data manipulation and extraction.
  • Excellent written and verbal communication skills.
  • Bachelor's degree in Computer Science or a related field.
  • Demonstrated ability to use AI to improve speed and quality in day-to-day workflow for relevant outputs.
  • Strong track record of critically evaluating and verifying AI-assisted work through testing, source-checking, data validation, or peer review.
  • High integrity and ownership, including protecting sensitive data and remaining accountable for final decisions and deliverables.
  • Experience in adtech is preferred.
  • Experience implementing data governance practices, including data quality, metadata management, and access controls, is preferred.
  • Strong understanding of privacy-by-design principles and handling sensitive or regulated data is preferred.
  • Familiarity with data table formats like Apache Iceberg or Delta is preferred.
  • Previous experience building out a Data Engineering function is preferred.
  • Proven experience working closely with Data Science teams on machine learning pipelines is preferred.

Benefits

  • Base salary range of $123,696 to $254,667 USD for US-based applicants.
  • Eligible for equity.
  • Remote-friendly working model noted via PinFlex.
  • Relocation assistance is not provided.
  • Additional culture and benefits information is available through Pinterest Life.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Lead Data Engineer

Media.Monks 5K-10K Media

Monks Technology Services is seeking a fully remote Lead Data Engineer for a fixed-term contract to lead the modernization of legacy Python/Spark ETL pipelines into dbt Core models on Databricks supporting analytics and business intelligence.

Apache Spark CI/CD Databricks Git GitHub Actions Python SQL
17 hours, 35 minutes ago

Staff Data Engineer

Webflow 251-1K Internet Software & Services

Webflow is hiring a Staff Data Engineer to lead data platform initiatives involving its data lake, event instrumentation, data quality, governance, and cloud infrastructure for modern web marketing operations.

Apache Airflow Apache Spark CI/CD Kafka SQL
17 hours, 35 minutes ago

Data Engineer [Zeal]

Livefront 11-50 Internet Software & Services

Zeal, now part of Livefront, is hiring a Data Engineer across its U.S. hubs to build reliable data pipelines and architectures that support technology consulting engagements for Fortune 1000 companies.

Apache Airflow CI/CD Databricks dbt GCP Git Kafka Oracle PostgreSQL Power BI Python RabbitMQ Snowflake SQL SQL Server Tableau
17 hours, 50 minutes ago

Sr Data Ops Engineer

Coderio 51-250 Internet Software & Services

Coderio is hiring a DataOps Engineer and Technical Referent to work with international customers on designing, implementing, and operating scalable cloud data infrastructure and automation solutions.

AWS Bash dbt Docker GitHub Python SQL Terraform
18 hours, 20 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers