Gather AI

Gather AI

Gather AI specializes in providing autonomous drone technology for inventory management in warehouses, enabling businesses to enhance inventory accuracy, reduce costs, and significantly improve operational productivity.

Air Freight & Logistics
11-50
$17M raised

Description

  • Build incremental extraction pipelines from production PostgreSQL into analytical warehouse layers.
  • Develop and maintain shared dbt models, dimensions, metric components, and serving tables.
  • Implement consistent metrics in the semantic layer across products and dashboards.
  • Add data tests, freshness checks, alerts, and safe backfill processes.
  • Apply tenant isolation and access-control patterns to data models and pipelines.
  • Maintain lineage linking metrics to source records, drone images, and video.
  • Validate incoming WMS data with the integration team against agreed contracts.
  • Document models and register them in the data catalog.
  • Deploy pipelines through CI/CD, use AI-assisted development, and participate in on-call.
  • Expand the shared data foundation across Drone, MHE Vision, and 3D case-counting products.

Requirements

  • 2–5 years of experience building and operating production data pipelines.
  • Degree in Computer Science or equivalent practical experience.
  • Strong SQL, including joins, window functions, CTEs, and dimensional modeling; PostgreSQL preferred.
  • Hands-on experience with dbt, Snowflake Dynamic Tables, Databricks Lakeflow Declarative Pipelines, or equivalent.
  • Production Python development with tests and code review experience.
  • Experience with Airflow, Dagster, Databricks Lakeflow Jobs, Snowflake Tasks, or equivalent orchestration tools.
  • Experience moving operational database data into a warehouse or lakehouse such as Snowflake or Databricks.
  • Production experience with Azure or another major cloud, plus object storage, Git, CI/CD, and Docker/Kubernetes.
  • Experience with data testing, pipeline failure handling, reruns, and backfills.
  • Clear written and spoken English, strong documentation skills, initiative, and ownership.
  • Experience with CDC or streaming tools such as Debezium, Fivetran, Kafka, or Event Hubs preferred.
  • Experience with PySpark, Snowpark, Terraform, semantic layers, data catalogs, multi-tenant platforms, image/video or sensor data, logistics, warehousing, or robotics preferred.

Benefits

  • Fully remote role on an India-based team.
  • Opportunity to build a data platform from the ground up in the robotics and warehouse intelligence industry.
  • Cross-functional collaboration with Full Stack, ML, Product, Security, and Customer Success teams.
  • Growth toward end-to-end ownership of a data domain, including design, quality, and on-call responsibilities.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Data Engineer

phData 251-1K IT Services

phData is seeking a Data Engineer in India to design, build, and productionize secure, scalable cloud data solutions for enterprise clients while collaborating with global delivery teams.

Apache Airflow Apache Spark AWS Azure Cassandra Databricks dbt Elasticsearch GCP Hadoop HDFS Kafka Luigi Python REST API Scala Snowflake SQL
1 hour, 26 minutes ago

Senior Data Engineer (US)

Firstup 251-1K Professional Services

Firstup is hiring a Senior Data Engineer to build scalable, reliable data pipelines and models while advancing GenAI-powered data interfaces and the company’s data architecture.

Apache Airflow Apache Spark AWS Databricks Generative AI Kafka Metabase Python React Ruby on Rails Snowflake SQL Trino
1 day, 1 hour ago

[job-31944] Senior Data Developer, Brasil.

CI&T 5K-10K Internet Software & Services

A CI&T busca um(a) Engenheiro(a) de Dados Sênior para construir e evoluir uma plataforma de dados Azure de grande escala, estruturando pipelines e análises confiáveis para apoiar decisões estratégicas do negócio.

Apache Spark Azure CI/CD Metabase Power BI Python SQL
1 day, 1 hour ago

Forward Deployed Engineer

Binagora 11-50 Internet Software & Services

Binagora is seeking a Forward Deployed Engineer to deploy and adapt data platforms within client environments while migrating infrastructure to GCP and building scalable pipelines for an influence measurement platform.

Apache Airflow Apache Spark Azure GCP Python Vertex AI
1 day, 2 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers