Lead / Principal Data Engineer - DBT & Databricks

5 hours, 22 minutes ago
Lead
Software Development
Coforge

Coforge

Coforge is a global digital services provider specializing in transforming businesses through technology and industry expertise, powering growth with innovative solutions and platforms.

IT Services
10K-50K
Founded 1992

Description

  • Design, build, and maintain DBT models on Azure Databricks using Delta Lake, Unity Catalog, and Medallion Architecture principles.
  • Deliver curated Gold-layer data domains to production standards in regulated enterprise environments.
  • Use AI-assisted development tools for model generation, refactoring, documentation, debugging, and engineering acceleration.
  • Establish DBT testing strategies, including schema validation, data quality testing, unit testing, and automated quality controls.
  • Work in CI/CD-driven delivery environments using pull request workflows and automated quality gates.
  • Curate datasets, semantic definitions, and business terminology for Databricks Genie and natural-language analytics initiatives.
  • Define and enforce data contracts, domain ownership, and governance standards.
  • Implement data observability, monitoring, and root-cause analysis practices across data platforms.
  • Drive modernization and migration initiatives through AI-assisted data engineering approaches.
  • Contribute to engineering playbooks, standards, and enablement programs for DBT, Databricks, and AI-powered development.

Requirements

  • 5+ years of experience in Data Engineering.
  • At least 2 years of hands-on experience building and operating production-grade DBT projects, including models, tests, macros, snapshots, and large-scale refactoring.
  • Strong SQL skills and solid Python programming experience.
  • Hands-on experience with Azure Databricks, including Delta Lake, Unity Catalog, and Medallion Architecture.
  • Strong knowledge of DBT testing frameworks, data validation, and quality assurance best practices.
  • Familiarity with dbt-expectations, Great Expectations, or Delta Live Tables expectations.
  • Experience working with Git-based pull request workflows, code reviews, and CI/CD deployment pipelines.
  • Experience using AI-powered development tools such as Claude Code, GitHub Copilot, Cursor, or similar solutions.
  • Ability to independently own data domains and collaborate directly with business stakeholders.
  • Strong communication, analytical, and problem-solving skills.
  • Experience curating Databricks Genie environments or building semantic layers for natural-language analytics (preferred).
  • Experience within Insurance or Financial Services domains (preferred).
  • Experience working with policy, claims, exposure, or regulatory reporting data (preferred).
  • Hands-on experience with PySpark for large-scale data transformations (preferred).
  • Experience with Databricks Workflows, Airflow, or similar orchestration platforms (preferred).
  • DBT and/or Databricks certifications (preferred).
  • Experience implementing AI-driven modernization and migration programs (preferred).

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

[Job-31034] Engenheiro de Dados (Azure/Databrics) Sênior , Brazil

CI&T 5K-10K Internet Software & Services

CI&T is hiring a Data Engineer to ensure data model integrity, consistency, efficiency, and governance across an Azure-based ecosystem.

Apache Spark Azure CI/CD Databricks GitHub MLflow Python SQL
5 hours, 37 minutes ago

Head of Fund Analytics & Automation — Credit Fund

Xsolla 251-1K Internet Software & Services

Xsolla is hiring a Head of Fund Analytics & Automation to build and own the data, automation, and reporting stack for its Credit Fund, supporting underwriting, investor reporting, and portfolio operations.

Apache Airflow PostgreSQL Power BI Python React SQL Tableau TypeScript
5 hours, 52 minutes ago

Associate Data Engineer

Abacus Insights 51-250 Insurance

Abacus Insights is hiring an Associate Data Engineer to support healthcare data distribution and client delivery on its cloud-based data management platform.

Apache Spark AWS Azure Databricks Snowflake SQL
5 hours, 52 minutes ago

[Job - 31058] Senior Data Developer, Brazil

CI&T 5K-10K Internet Software & Services

CI&T busca um(a) Desenvolvedor(a) de Dados Sênior para atuar na estratégia de dados de um cliente de mídia e entretenimento, integrando sistemas críticos e entregando uma visão unificada e confiável das informações corporativas.

CRM ERP GCP Salesforce SQL
5 hours, 52 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers