Data Engineer (Azure) - Remote, Latin America

1 month ago
Full-time
Mid Level
Software Development
Bluelight Consulting

Bluelight Consulting

Bluelight Consulting is a leading Nearshore Software Development company that provides access to highly skilled Nearshore Tech Talent in the US timezone. They offer services such as Nearshore Staffing, Cloud consulting, Cloud migration strategies, and ...

Internet Software & Services
11-50
Founded 2015

Description

  • Develop and maintain ETL processes using Python (PySpark) in Azure Synapse Analytics notebooks and pipelines.
  • Design and build data warehouse structures using star schemas, facts, and dimensions in an MPP SQL pool.
  • Extract and integrate data from REST APIs, SQL database tables, and CSV files.
  • Design and optimize Azure Synapse notebooks and pipelines for scalability and performance.
  • Contribute to data fabric capabilities including data lakes, lakehouses, delta lakes, and data cataloging.
  • Collaborate with data architects to create data models and schemas that align with business needs.
  • Implement data quality checks and validation processes to ensure accurate and consistent data.
  • Identify and resolve performance bottlenecks and troubleshoot ETL jobs to meet SLAs.
  • Maintain documentation for ETL processes, data flows, and transformations.
  • Work with cross-functional teams to understand requirements and support data-related initiatives.
  • Ensure data security, governance, and privacy compliance.

Requirements

  • Bachelor’s degree in Computer Science, Information Technology, or a related field, or equivalent work experience.
  • Certifications related to data engineering or data science, such as Azure Data Engineer, are a plus.
  • Proven experience in ETL data engineering using Python (PySpark).
  • Experience extracting, transforming, and loading data from REST APIs, SQL database tables, and CSV files.
  • Proficiency with Azure Synapse Analytics, including Notebooks, Pipelines, Linked Services, and Azure Key Vault.
  • Ability to write complex SQL queries and optimize query performance.
  • Experience with both SparkSQL and MS SQL.
  • Knowledge of data integration best practices and tools.
  • Experience with version control systems such as Git and Azure DevOps.
  • Strong problem-solving and analytical skills with attention to detail.
  • Excellent verbal and written communication skills and ability to work in a team with shifting priorities.
  • Familiarity with big data technologies, machine learning, and data analysis is preferred.
  • Experience with data visualization tools such as Power BI or Tableau is a plus.
  • Experience with Agile methodologies is a plus.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Integrator

Clarity Innovations 51-250 Internet Software & Services

Clarity Innovations is hiring a Data Integrator to support AFSOC data initiatives by turning disparate data into actionable insights for mission decision-making within the national security and defense environment.

CI/CD DevSecOps
7 hours, 22 minutes ago

Software Engineer - X Data

x.ai 51-250 Internet Software & Services

SpaceXAI is hiring a Data Engineer for its X product engineering team to build and operate data infrastructure that supports product, business, and internal stakeholder decisions.

Apache Spark ClickHouse Flink GCP Go Java Kafka Python Rust Scala SQL Trino
7 hours, 52 minutes ago

Grupo QuintoAndar | Senior Data Engineer (Women Applicants Only)

QuintoAndar 1K-5K Real Estate

QuintoAndar Group is hiring a Data Engineer Specialist in its Brazil-based remote-first technology team to design and maintain scalable, secure data platforms and pipelines that support analytics and strategic decision-making.

Apache Airflow Apache Spark AWS CI/CD Databricks Grafana Hadoop Hive Kubernetes MapReduce Microservices Prometheus Python SQL Terraform Trino YAML
7 hours, 52 minutes ago

Senior Software Engineer - Data

Senior Software Engineer at ApartmentIQ, a profitable remote PropTech company, to evolve the shared data platform powering its multifamily data products and improve coverage, freshness, and insights.

Apache Airflow AWS dbt JSON Linear OpenSearch Python Ruby on Rails XML
7 hours, 52 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers