ParetoHealth

ParetoHealth

ParetoHealth revolutionizes employee health benefits with the largest captive in the nation, reducing costs, capping risks, and empowering businesses to control healthcare spending.

Insurance
51-250
Founded 2011

Description

  • Design, implement, and maintain scalable serverless data pipelines on AWS for reporting, analytics, and AI use cases.
  • Build and evolve data models and schemas for performant querying and downstream consumption.
  • Develop ETL/ELT workflows to ingest, cleanse, transform, and load data from internal, third-party, and streaming sources.
  • Partner with Product, Underwriting, Analytics, and AI teams to define data requirements, contracts, and SLAs.
  • Implement data quality controls, monitoring, and alerting for critical datasets and features.
  • Optimize serverless workloads for cost, performance, and scalability, including Athena tuning and S3 storage design.
  • Apply data engineering best practices, including version control, code review, CI/CD, and Infrastructure as Code.
  • Collaborate on feature stores and reusable data assets for experimentation and model deployment.
  • Troubleshoot pipeline issues and support production data workflows and model-driven applications.
  • Document data models, pipelines, and data contracts while promoting data literacy and self-service analytics.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, Mathematics, or related field, or equivalent practical experience.
  • 3+ years of experience in data engineering or software engineering roles focused on data pipelines and analytics platforms.
  • 2+ years of hands-on experience with AWS data services in a production environment.
  • Strong experience with AWS serverless/data services such as Lambda, Glue, Athena, S3, and Step Functions or similar orchestration tools.
  • Proficiency in Python and TypeScript or a similar language for data engineering.
  • Solid understanding of data modeling principles, including dimensional, normalized, wide-table, and event-driven designs.
  • Experience designing and operating data pipelines at scale, including batch and near-real-time ingestion.
  • Familiarity with SQL and query optimization in columnar data stores and engines such as Athena.
  • Knowledge of data quality, governance, and security practices, including handling sensitive healthcare and financial data.
  • Hands-on experience with Infrastructure as Code, preferably AWS CDK.
  • Ability to collaborate with Analytics and AI teams and translate model data needs into data solutions.
  • Strong communication skills and ability to explain complex data concepts clearly to business stakeholders.

Benefits

  • Fully paid medical, dental, and vision benefits.
  • Flexible PTO.
  • 401(k) company contribution.
  • Tuition reimbursement.
  • Professional development allowance.
  • Transportation allowance and daily parking reimbursement.
  • Engaging hybrid work environment.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Data Engineer

Nimble Gravity 51-250 IT Services

Nimble Gravity is seeking a Senior Data Engineer to lead the modernization of client data ecosystems into a Snowflake-based cloud platform supporting decentralized Data Mesh and AI-ready self-service capabilities.

AWS Azure CI/CD dbt Python Snowflake SQL
20 hours, 25 minutes ago

Senior Data Engineer and Analytics Lead

Open Data Jobs Professional Services

Peregrine Advisors is hiring a remote Data Engineer to provide technical continuity across three production data products for a federal agency, building reliable pipelines, governed data models, and supportable analytics solutions.

Apache Spark AWS CI/CD Databricks Python SQL
21 hours, 25 minutes ago

Senior Data Engineer

Babylist 251-1K Internet Software & Services

Babylist is hiring a senior Data Engineering individual contributor to build scalable data infrastructure and AI-native systems that automate pipeline generation, testing, and maintenance across its growing e-commerce platform.

Apache Airflow AWS dbt Python Snowflake
1 day, 20 hours ago

Senior Data Engineer (India)

Alimentiv 251-1K Professional Services

The Lead Data Engineer will architect and operationalize enterprise-scale data solutions for analytics and AI/ML initiatives, ensuring integrated, governed, and production-ready data platforms while providing technical leadership.

Apache Spark Azure CI/CD Databricks Git HIPAA JSON Looker Power BI REST API SQL Tableau Terraform
1 day, 20 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers