Data Engineer

1 hour, 37 minutes ago
Full-time
Mid Level
Software Development
ParetoHealth

ParetoHealth

ParetoHealth revolutionizes employee health benefits with the largest captive in the nation, reducing costs, capping risks, and empowering businesses to control healthcare spending.

Insurance
51-250
Founded 2011

Description

  • Design, implement, and maintain scalable serverless data pipelines on AWS for reporting, analytics, and AI use cases.
  • Build and evolve data models and schemas for performant querying and downstream consumption.
  • Develop ETL/ELT workflows to ingest, cleanse, transform, and load data from internal, third-party, and streaming sources.
  • Partner with Product, Underwriting, Analytics, and AI teams to define data requirements, contracts, and SLAs.
  • Implement data quality controls, monitoring, and alerting for critical datasets and features.
  • Optimize serverless workloads for cost, performance, and scalability, including Athena tuning and S3 storage design.
  • Apply data engineering best practices, including version control, code review, CI/CD, and Infrastructure as Code.
  • Collaborate on feature stores and reusable data assets for experimentation and model deployment.
  • Troubleshoot pipeline issues and support production data workflows and model-driven applications.
  • Document data models, pipelines, and data contracts while promoting data literacy and self-service analytics.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, Mathematics, or related field, or equivalent practical experience.
  • 3+ years of experience in data engineering or software engineering roles focused on data pipelines and analytics platforms.
  • 2+ years of hands-on experience with AWS data services in a production environment.
  • Strong experience with AWS serverless/data services such as Lambda, Glue, Athena, S3, and Step Functions or similar orchestration tools.
  • Proficiency in Python and TypeScript or a similar language for data engineering.
  • Solid understanding of data modeling principles, including dimensional, normalized, wide-table, and event-driven designs.
  • Experience designing and operating data pipelines at scale, including batch and near-real-time ingestion.
  • Familiarity with SQL and query optimization in columnar data stores and engines such as Athena.
  • Knowledge of data quality, governance, and security practices, including handling sensitive healthcare and financial data.
  • Hands-on experience with Infrastructure as Code, preferably AWS CDK.
  • Ability to collaborate with Analytics and AI teams and translate model data needs into data solutions.
  • Strong communication skills and ability to explain complex data concepts clearly to business stakeholders.

Benefits

  • Fully paid medical, dental, and vision benefits.
  • Flexible PTO.
  • 401(k) company contribution.
  • Tuition reimbursement.
  • Professional development allowance.
  • Transportation allowance and daily parking reimbursement.
  • Engaging hybrid work environment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer (6278)

Dan.com - a GoDaddy brand Internet Software & Services

itD is hiring a remote Data Engineer to build and maintain scalable data infrastructure for AI product analytics and topline metrics across product surfaces.

AWS Azure GCP Generative AI Machine Learning Python Snowflake SQL
37 minutes ago

Lead Data Manager/SAS programmer (Job 1425)

DLH 1K-5K Construction & Engineering

DLH is seeking a Lead Data Manager/SAS Programmer to lead data management and statistical programming for federally funded and commercially funded clinical research and epidemiological studies.

GCP Git HIPAA Python R SQL
37 minutes ago

Data Engineer, AWS Glue

NTT DATA 100K+ IT Services

NTT DATA is hiring a remote Data Engineer, AWS Glue to support a U.S. global client from LATAM, focusing on building and operating enterprise data ingestion and lake processing pipelines on AWS.

Apache Spark AWS CloudFormation Python SQL Terraform
37 minutes ago

Data Engineer with Airflow

Xebia 1K-5K Internet Software & Services

Xebia is hiring a Data Engineer to support a managed Apache Airflow platform, helping customer data teams adopt, integrate, and operate Airflow in production environments.

Apache Airflow AWS Azure Databricks dbt GCP Kubernetes Python Snowflake SQL Terraform
37 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers