phData

phData

phData specializes in providing comprehensive data engineering, data strategy, and machine learning services, offering end-to-end managed data solutions that facilitate data migrations, integrations, and the development of modern data products and appl...

IT Services
251-1K

Description

  • Develop end-to-end technical solutions and help deliver them into production.
  • Ensure solutions meet performance, security, scalability, and data integration needs.
  • Write, debug, and optimize SQL queries.
  • Create and deliver detailed presentations for clients and stakeholders.
  • Produce detailed solution documentation, including POCs, roadmaps, sequence diagrams, class hierarchies, and logical system views.
  • Collaborate in client-facing work and communicate clearly in written and verbal English.
  • Support the full software development lifecycle, including design, implementation, testing, and deployment.

Requirements

  • 4+ years of experience as a Software Engineer, Data Engineer, or Data Analyst.
  • Programming experience in Java, Python, and/or Scala.
  • Experience with core cloud data platforms such as Snowflake, AWS, Azure, Databricks, and GCP.
  • Strong SQL skills with the ability to write, debug, and optimize queries.
  • Client-facing English written and verbal communication experience.
  • Ability to create and deliver detailed presentations.
  • Experience creating detailed solution documentation.
  • Bachelor’s degree in Computer Science or a related field.
  • Preferred experience in production environments on Snowflake, AWS, Azure, GCP, Hadoop, or Databricks.
  • Preferred experience with cloud and distributed data storage such as S3, ADLS, HDFS, GCS, Kudu, Elasticsearch/Solr, Cassandra, or other NoSQL systems.
  • Preferred experience with data integration tools and technologies such as Spark, Kafka, StreamSets, Matillion, Fivetran, NiFi, AWS DMS, Azure Data Factory, Informatica IICS, or Google DataProc.
  • Preferred experience working with multiple data sources such as queues, relational databases, files, search, and APIs.
  • Preferred full SDLC experience including design, documentation, implementation, testing, and deployment.
  • Preferred experience with automated data transformation and curation tools such as dbt, Spark, Spark streaming, and automated pipelines.
  • Preferred experience with workflow management and orchestration tools such as Airflow, AWS Managed Airflow, Luigi, or NiFi.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer

CodeRoad 51-250 Internet Software & Services

Coderoad is hiring a Data Engineer to build and improve cloud-based data pipelines and analytics infrastructure for real-world software projects.

Apache Spark Databricks GCP MySQL Python Scala
19 hours, 3 minutes ago

Data Engineer

Total Performance Consulting 51-250 Internet Software & Services

A fully remote Data Engineer role at Perform in Latin America focused on building scalable data pipelines and analytics infrastructure across the organization.

Apache Airflow AWS dbt Looker Machine Learning Python Snowflake SQL Terraform
19 hours, 3 minutes ago

Freelance Data Scraping Engineer (Python)

Mindrift.ai: Be the “I” in AI Internet Software & Services

Mindrift is hiring a freelance Python Data Scraping Engineer for the Tendem project to build and maintain data extraction workflows for real-world web data use cases.

AJAX GitHub HTML JavaScript JSON LLM Python Selenium
20 hours, 3 minutes ago

Data Platform Engineer (GCP)

Innodata 1K-5K IT Services

Innodata is hiring a hands-on GCP/Data Platform Engineer to support and improve large-scale production data platforms on Google Cloud, with a focus on reliability, troubleshooting, and operational enhancement.

Apache Airflow CI/CD GCP Generative AI LLM Python SQL Terraform Vertex AI
1 day, 19 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers