ParetoHealth

ParetoHealth

ParetoHealth revolutionizes employee health benefits with the largest captive in the nation, reducing costs, capping risks, and empowering businesses to control healthcare spending.

Insurance
51-250
Founded 2011

Description

  • Design, implement, and maintain scalable serverless data pipelines on AWS for reporting, analytics, and AI use cases.
  • Build and evolve data models and schemas for performant querying and downstream consumption.
  • Develop ETL/ELT workflows to ingest, cleanse, transform, and load data from internal, third-party, and streaming sources.
  • Partner with Product, Underwriting, Analytics, and AI teams to define data requirements, contracts, and SLAs.
  • Implement data quality controls, monitoring, and alerting for critical datasets and features.
  • Optimize serverless workloads for cost, performance, and scalability, including Athena tuning and S3 storage design.
  • Apply data engineering best practices, including version control, code review, CI/CD, and Infrastructure as Code.
  • Collaborate on feature stores and reusable data assets for experimentation and model deployment.
  • Troubleshoot pipeline issues and support production data workflows and model-driven applications.
  • Document data models, pipelines, and data contracts while promoting data literacy and self-service analytics.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, Mathematics, or related field, or equivalent practical experience.
  • 3+ years of experience in data engineering or software engineering roles focused on data pipelines and analytics platforms.
  • 2+ years of hands-on experience with AWS data services in a production environment.
  • Strong experience with AWS serverless/data services such as Lambda, Glue, Athena, S3, and Step Functions or similar orchestration tools.
  • Proficiency in Python and TypeScript or a similar language for data engineering.
  • Solid understanding of data modeling principles, including dimensional, normalized, wide-table, and event-driven designs.
  • Experience designing and operating data pipelines at scale, including batch and near-real-time ingestion.
  • Familiarity with SQL and query optimization in columnar data stores and engines such as Athena.
  • Knowledge of data quality, governance, and security practices, including handling sensitive healthcare and financial data.
  • Hands-on experience with Infrastructure as Code, preferably AWS CDK.
  • Ability to collaborate with Analytics and AI teams and translate model data needs into data solutions.
  • Strong communication skills and ability to explain complex data concepts clearly to business stakeholders.

Benefits

  • Fully paid medical, dental, and vision benefits.
  • Flexible PTO.
  • 401(k) company contribution.
  • Tuition reimbursement.
  • Professional development allowance.
  • Transportation allowance and daily parking reimbursement.
  • Engaging hybrid work environment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

[Job - 31429] Data Engineer Specialist

CI&T 5K-10K Internet Software & Services

CI&T is seeking a Data Engineer to own the Foundation Pod’s end-to-end data platform, building reliable pipelines, governed models, and analytics solutions for enterprise transformation initiatives.

Azure CI/CD Power BI Terraform
11 hours, 2 minutes ago

Senior Software Engineer II (Data Pipelines)

Wpromote 251-1K Media

Wpromote is seeking a Senior Engineer to lead the development, automation, and optimization of scalable data pipelines supporting Polaris, its marketing intelligence and operational data platform.

Agile Apache Airflow CI/CD dbt GCP Git GitHub Actions Grafana Hive Kubernetes Looker PostgreSQL Prometheus Python Scrum SQL Terraform
11 hours, 2 minutes ago

Director, Data Engineering - GCP

TTEC Digital 1K-5K Professional Services

TTEC Digital is seeking a Director of Data Engineering (GCP) to lead its Google Cloud data engineering practice and build scalable data platforms supporting analytics, AI/ML, governance, and digital transformation initiatives.

Agile Apache Airflow Apache Spark CI/CD GCP Generative AI Machine Learning Python SQL Terraform Vertex AI
11 hours, 2 minutes ago

Mid/Senior-level GoLang Data Engineer (EU)

Inca Digital 11-50 Internet Software & Services

Inca Digital is seeking a Golang Backend Engineer to build and operate data infrastructure that transforms blockchain, financial, social, and web data into intelligence products for digital-asset risk, surveillance, and investigations.

AWS CI/CD Docker Git Go gRPC Kafka Kubernetes Microservices MongoDB NLP PostgreSQL RabbitMQ Redis REST API SQL Terraform
11 hours, 2 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers