CI&T

CI&T

CI&T is a global digital technology agency empowering agile growth for leading companies through advanced technologies with a team of 2000 experts worldwide.

Internet Software & Services
5K-10K
Founded 1995

Description

  • Design, build, and maintain ETL/ELT pipelines for a modern data lake architecture.
  • Develop and optimize distributed data processing workflows using Python and PySpark.
  • Implement partitioning strategies to balance query performance, storage cost, and maintainability.
  • Write, optimize, and translate complex SQL queries for data transformation and modeling.
  • Migrate and modernize legacy RDBMS data pipelines into cloud-native analytics environments.
  • Contribute to reusable internal code libraries using object-oriented programming principles.
  • Work with AWS services such as Glue, Athena, Redshift, S3, Lambda, and EventBridge.
  • Collaborate with infrastructure and DevOps teams to provision data resources with IaC tools.
  • Monitor pipeline health, performance, and reliability using CloudWatch and observability tools.
  • Partner with analysts, scientists, and business stakeholders to turn requirements into scalable technical solutions.

Requirements

  • Experience building ETL processes and data pipelines with AWS.
  • Strong proficiency in Python as the primary programming language.
  • Hands-on experience writing and optimizing PySpark code for distributed data processing.
  • Strong SQL knowledge, including CTEs, window functions, aggregations, and conditional expressions.
  • Experience translating workloads from legacy RDBMS platforms to cloud environments.
  • Hands-on experience with AWS Glue, Athena, and Redshift.
  • Solid understanding of data lake architectures and partitioning strategies.
  • Working knowledge of object-oriented programming and reusable code libraries.
  • Comfort with Git, shell scripts, and Linux environments.
  • Familiarity with observability, monitoring, and metric tracking practices.
  • English at an advanced/fluent level.
  • Preferred experience with Delta Lake or Apache Iceberg.
  • Preferred knowledge of TypeScript.
  • Preferred experience with IaC tools such as CloudFormation, CDK, or Terraform.
  • Preferred familiarity with AWS services such as SageMaker AI, ECS, RDS, DynamoDB, IAM, or EventBridge.
  • Preferred experience with Pandas for data manipulation and analysis.
  • Preferred exposure to machine learning workflows or AI-driven data initiatives.

Benefits

  • Health and dental insurance.
  • Meal and food allowance.
  • Childcare assistance.
  • Extended paternity leave.
  • Access to gyms and wellness professionals through Wellhub (Gympass) and TotalPass.
  • Profit sharing and results participation (PLR).
  • Life insurance.
  • Continuous learning through CI&T University and partnerships with online learning platforms.
  • Discount club benefits.
  • Free online platform for physical, mental, and overall well-being.
  • Pregnancy and responsible parenting course.
  • Language learning platform.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer IV (6249)

Dan.com - a GoDaddy brand Internet Software & Services

itD is hiring a Data Engineer IV to build and optimize AWS-based data infrastructure and pipelines that support AI infrastructure and inference initiatives in a remote role.

AWS
13 hours, 39 minutes ago

Senior Data Engineer (GCP)

Xebia 1K-5K Internet Software & Services

Xebia is hiring a data engineer to support a BigQuery migration by building and maintaining scalable data pipelines for production reporting, analytics, and business intelligence solutions.

Apache Airflow CI/CD Databricks dbt Git Python Snowflake SQL
13 hours, 39 minutes ago

Staff Data Engineer (Coupang Pay)

Coupang 1K-5K Internet Software & Services

쿠팡페이 핀테크 데이터 플랫폼 팀의 Staff Data Engineer로서 실시간 결제·주문 데이터를 기반으로 데이터 파이프라인과 분석 인프라를 구축·고도화하는 역할입니다.

Apache Spark AWS ClickHouse Flink Hive Java Presto Python Scala SQL Tableau
13 hours, 54 minutes ago

Data Migration Engineer (SQL Server / TSQL)

Cresteo 51-200 information technology & services

Cresteo is hiring a Data Migration Engineer to own end-to-end SQL Server data migrations from legacy schemas into reconciled, production-ready data for US-based clients and international teams.

JSON SQL Server
13 hours, 54 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers