RevStar

RevStar

RevStar Consulting is a client-centric cloud consulting firm that specializes in developing modern, user-focused, cloud-native web and mobile applications. They offer custom integrations, implementations, and solutions using the latest cloud technologi...

Internet Software & Services
51-250
Founded 2009

Description

  • Develop and optimize data pipelines using Apache Spark and Delta Lake within Databricks.
  • Implement ETL/ELT workflows for data ingestion, transformation, and storage.
  • Design scalable Lakehouse architecture solutions across structured and unstructured data sources.
  • Integrate Databricks with cloud storage platforms such as Azure Data Lake, AWS S3, and Google Cloud Storage.
  • Optimize Spark jobs for scalability, cost efficiency, and low latency.
  • Implement monitoring, alerting, automated validation, and data quality processes.
  • Support ML model training and deployment within Databricks using MLflow for tracking and versioning.
  • Collaborate with data scientists, ML engineers, architects, and business stakeholders to deliver aligned solutions.
  • Implement feature engineering pipelines and help integrate models into production environments.
  • Ensure data security, access control, governance, compliance, and documentation best practices.

Requirements

  • 3+ years of hands-on experience in data engineering with big data processing and cloud-native architectures.
  • 2+ years of hands-on experience with Databricks, including Apache Spark, Delta Lake, and MLflow.
  • Databricks Certified Data Engineer Associate or higher certification is mandatory.
  • Proficiency in Python, SQL, and Spark-based frameworks.
  • Experience developing and optimizing large-scale ETL/ELT pipelines.
  • Strong understanding of Lakehouse architecture and cloud-agnostic data solutions.
  • Familiarity with CI/CD pipelines and Infrastructure-as-Code tools for Databricks, such as Terraform and Databricks CLI.
  • Knowledge of data governance, security, and compliance best practices.
  • Experience working in Agile environments and following DevOps/MLOps best practices.
  • Preferred: additional Databricks certifications, such as Databricks Certified Machine Learning Associate.
  • Preferred: experience with real-time streaming tools such as Kafka, Kinesis, or Event Hub.
  • Preferred: familiarity with orchestration tools such as Apache Airflow or Prefect.
  • Preferred: background in AI/ML integration within Databricks, including feature engineering and model deployment.
  • Preferred: experience in client-facing roles or consulting environments.

Benefits

  • Paid time off.
  • Remote-first working environment.
  • Comprehensive health coverage including medical, dental, and vision.
  • 401(k) retirement plan.
  • Annual learning and development stipend for conferences, certifications, or courses.
  • Peer mentorship and coaching.
  • Professional growth opportunities with exposure to AWS GenAI, data, and cloud technologies.
  • Company outings and volunteer opportunities.
  • Collaborative, innovative culture.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer IV (6329)

Dan.com - a GoDaddy brand Internet Software & Services

itD is seeking a remote Data Engineer IV for a 12-month engagement to scale data infrastructure, enable AI-driven analytics, and deliver customer experience insights for its Post Sales Organization.

Apache Airflow Apache Spark CI/CD Feature Engineering Generative AI Git Google Analytics Kafka Python Snowflake SQL Tableau
7 hours, 52 minutes ago

(1573) Data Engineer (PostgreSQL & Databricks)

Nortal 251-1K Internet Software & Services

Nortal is building a LATAM remote talent network for Data Engineers to design, develop, maintain, and support scalable data pipelines and customer-facing applications using Databricks, Postgres, and SQL Server.

Agile AWS CI/CD Databricks DevSecOps Microservices PostgreSQL Scrum Serverless SQL Server
8 hours, 7 minutes ago

Data Platform Engineer

Figma 1K-5K Internet Software & Services

Figma is hiring a Data Platform Engineer to build foundational data and infrastructure systems that support analytics, machine learning, and AI-driven product experiences.

CI/CD Machine Learning Python
8 hours, 51 minutes ago

Marketing Data Engineer - Ad Tech (US hours)

VirtuHire 11-50 Staffing / Recruiting

The Marketing Data Engineer will own the production data pipelines and data-quality foundation that support dashboards, pacing outputs, and recurring reporting for a U.S.-based client.

Python Snowflake SQL
9 hours, 21 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers