CI&T

CI&T

CI&T is a global digital technology agency empowering agile growth for leading companies through advanced technologies with a team of 2000 experts worldwide.

Internet Software & Services
5K-10K
Founded 1995

Description

  • Define the target solution architecture for data structure, processing, storage, consumption, and governance in collaboration with the project architecture team.
  • Conduct structured assessments of existing ETL/ELT, reporting, and integration workloads and map cross-system dependencies.
  • Recommend and justify rationalization decisions for the new solution, including migrate, refactor, eliminate, or defer options.
  • Identify technical risks, mitigation plans, and quick wins for the modernization journey.
  • Define and implement data quality gates across pipelines, including completeness, freshness, accuracy, and lineage.
  • Establish governance and security policies with Unity Catalog, including access control, RBAC/ABAC, sensitive data masking, lineage, and auditability.
  • Define pipeline orchestration, deployment, and observability strategies using Databricks Workflows, Airflow, or equivalent tools.
  • Design platform operations, including monitoring, alerts, failure handling, autoscaling, and cost optimization.
  • Serve as a technical reference for the data team, supporting solution design, code reviews, and key technical decisions.
  • Align with business, corporate architecture, and infrastructure teams to ensure enterprise standards and strategic goals are met.
  • Influence roadmap and prioritization based on business value, technical effort, and risk.
  • Produce and maintain architecture artifacts such as diagrams, ADRs, runbooks, dependency maps, and guidelines, and lead knowledge transfer sessions.

Requirements

  • 5+ years of experience in Data Engineering or Data Architecture, including 2–3 years in a Data Architect or equivalent technical leadership role.
  • Proven experience leading architecture for at least one medium-to-large data modernization or migration program.
  • Hands-on experience with relational and non-relational data models, including MongoDB, DynamoDB, PostgreSQL, and similar technologies.
  • Strong knowledge of performance best practices such as partitioning, Z-order, caching, and job optimization.
  • Ability to define standards for data platforms, including cluster strategy, aging, cost control, security, and governance.
  • Practical experience with AWS tools for large-scale data storage and processing.
  • Ability to design architectures that consider security, scalability, FinOps, and integration with the client ecosystem.
  • Experience defining and implementing data governance frameworks, including data domains, ownership, access policies, and cataloging.
  • Experience defining and measuring data SLAs/SLOs, quality checks, and monitoring.
  • Proven experience migrating legacy pipelines and reports such as SQL Server, SSIS, and traditional EDW environments to modern platforms.
  • Ability to read, interpret, and assess legacy code such as T-SQL, SSIS XML/dtsx, and PySpark to guide migration or refactoring strategies.
  • Experience with pipeline orchestration tools such as Apache Airflow and Databricks Workflows.
  • Experience with CI/CD practices for data, including IaC, environment promotion, and code/schema versioning.
  • Strong communication skills for both executive and technical audiences, including architecture presentation, trade-offs, and risk communication.
  • Experience working in agile teams and collaborating with Product, Technology, and Business stakeholders.
  • Ability to influence decisions and lead complex technical discussions in ambiguous environments.
  • Experience with Lakebridge (Databricks Labs) or other SQL migration automation tools is a plus.
  • Familiarity with AI-assisted code conversion approaches, such as LLMs for T-SQL to Spark SQL or SSIS to notebooks, is a plus.
  • Background with PySpark, Sqoop, or ingestion pipelines in Hadoop/EMR ecosystems is a plus.
  • Experience comparing orchestration patterns across Apache Airflow and Databricks Workflows is a plus.
  • Experience with knowledge graphs or metadata management for dependency tracking is a plus.
  • Relevant certifications such as Databricks Lakehouse/Data Engineer or AWS Data Analytics Specialty are a plus.
  • Prior experience in the education sector or in environments with multiple academic and administrative sources such as CRM, LMS, and ERP is a plus.

Benefits

  • Health and dental insurance.
  • Meal and food allowance.
  • Childcare assistance.
  • Extended parental leave.
  • Gym and wellness partnerships through Wellhub (Gympass) and TotalPass.
  • Profit sharing program (PLR).
  • Life insurance.
  • Continuous learning platform through CI&T University.
  • Discount club.
  • Free online platform focused on physical, mental, and overall well-being.
  • Pregnancy and responsible parenting course.
  • Partnerships with online learning platforms.
  • Language learning platform.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Sr. Manager, Data Engineering

Carrot Fertility 251-1K Insurance

Carrot is hiring a Sr. Manager, Data Engineering to lead its BI data engineering function, shaping the data platform and team that power analytics and reporting across a regulated global healthcare business.

Apache Airflow AWS CI/CD dbt Git HIPAA Prefect Python Snowflake SQL
2 days, 15 hours ago

Staff Technical Data Architect, Integrations

Lighthouse Hotels, Restaurants & Leisure

Lighthouse is hiring a Staff Technical Data Architect for its Integrations team in Spain to shape the architecture of large-scale hospitality data systems that power its AI-driven platform.

Apache Airflow Databricks GCP Kubernetes Python Snowflake
1 week, 1 day ago

Senior Data Architect - Databricks Platform

Caylent 251-1K IT Services

Caylent is seeking a Databricks Architect to design and harden enterprise-scale data platforms for client engagements, creating the governance and automation needed to move modern data capabilities into production.

Apache Spark AWS CI/CD Databricks Terraform
1 week, 1 day ago

Principal Data Architect — Databricks Enablement

Caylent 251-1K IT Services

Caylent is seeking a Principal Data Architect to lead enterprise Databricks engagements, guiding clients in building governable data platforms and delivering their first priority use cases from design through production.

Apache Spark AWS Databricks Generative AI Machine Learning Matplotlib MLOps NumPy Pandas Python Terraform
1 week, 2 days ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers