Bounteous

Bounteous

Bounteous collaborates with ambitious brands across various industries, including finance and healthcare, to drive digital transformation by creating innovative and seamless customer experiences through advanced technologies such as AI, cloud, and data...

Internet Software & Services
1K-5K
Founded 2003

Description

  • Architect and lead implementation of an enterprise lakehouse on Databricks across one or more major cloud providers.
  • Design batch and streaming data pipelines using PySpark, Spark SQL, Structured Streaming, and Delta Live Tables.
  • Establish platform patterns for ingestion from operational systems, event streams, and third-party APIs.
  • Define and enforce standards for data modeling, CI/CD, code quality, testing, observability, and cost optimization.
  • Lead the governance strategy with Unity Catalog, including access control, lineage, auditability, and PII handling.
  • Optimize Spark workloads for performance and cost through cluster sizing, Photon, autoscaling, file layout, Z-ordering, caching, and query tuning.
  • Partner with ML engineers and data scientists to operationalize models using MLflow, feature stores, and model serving.
  • Own cloud infrastructure components including networking, IAM, secrets, encryption, and Terraform/IaC for Databricks and supporting services.
  • Mentor data engineers and lead architecture reviews, code reviews, and technical design sessions.
  • Work with stakeholders across analytics, product, and finance to shape the data platform roadmap.

Requirements

  • 8+ years of data engineering experience, including 4+ years building production workloads on Databricks.
  • Deep expertise in Apache Spark, including PySpark, Spark SQL, performance tuning, partitioning strategy, and the Catalyst/Photon execution model.
  • Strong hands-on experience with Delta Lake, Unity Catalog, Databricks Workflows, and Delta Live Tables.
  • Production experience on at least one major cloud platform: AWS, Azure, or GCP.
  • Experience with cloud networking, IAM, storage services such as S3, ADLS, or GCS, and compute primitives.
  • Proficiency in Python and SQL; Scala experience is a plus.
  • Experience designing medallion (bronze/silver/gold) architectures and dimensional models for analytics.
  • Strong CI/CD and DevOps practices, including Git, Terraform, Databricks Asset Bundles or dbx, and automated testing of data pipelines.
  • Track record of leading technical projects end to end and mentoring engineers.
  • Excellent written and verbal communication skills for aligning technical and business stakeholders.
  • Must be legally eligible to work in Canada.
  • Employer is willing to sponsor eligible candidates for employment visas.

Benefits

  • Salary range of $115,000 to $150,000 per year.
  • Remote work arrangement.
  • Visa sponsorship available for eligible candidates.
  • Equal opportunity employer with accommodations provided during the hiring process in Ontario.
  • Opportunity to work on enterprise-scale Databricks, cloud, data, and AI initiatives.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Data Engineer

phData 251-1K IT Services

phData is hiring a Data Engineer in India to design and deliver secure, scalable cloud data solutions for enterprise clients as part of its remote-first consulting team.

Apache Airflow Apache Spark AWS Azure Cassandra Databricks dbt Elasticsearch GCP Hadoop HDFS Kafka Luigi Python Scala Snowflake Solr SQL
19 hours, 53 minutes ago

Principal Solution Architect- Managed Services

phData 251-1K IT Services

phData is hiring a Principal Solutions Architect for its Managed Services EPO team in India to lead the design and delivery of large-scale cloud data platforms for global enterprise clients.

Apache Airflow Apache Spark AWS Azure Bitbucket Cassandra CI/CD CloudFormation Databricks dbt GCP GitHub Kafka Luigi Machine Learning Python Snowflake SQL Terraform
20 hours, 8 minutes ago

Manager, Software Engineering - Data Platform

Figma 1K-5K Internet Software & Services

Figma is hiring a Data Platform engineering leader to scale the systems that power core business metrics, product analytics, and applied AI and experimentation across the company.

Machine Learning
20 hours, 8 minutes ago

Lead Bioinformatician/Engineer (Pipelines and Infrastructure)

Natera 1K-5K Pharmaceuticals

Natera is hiring a Lead Bioinformatics Engineer to own pipeline and infrastructure engineering for a research informatics team developing cfDNA screening workflows in women’s and organ health.

AWS Git Linux Python Terraform
20 hours, 23 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers