RevStar

RevStar

RevStar Consulting is a client-centric cloud consulting firm that specializes in developing modern, user-focused, cloud-native web and mobile applications. They offer custom integrations, implementations, and solutions using the latest cloud technologi...

Internet Software & Services
51-250
Founded 2009

Description

  • Design, build, and optimize scalable ETL/ELT pipelines using Apache Spark and Delta Lake across multi-cloud environments.
  • Implement Lakehouse architectures that process structured and unstructured enterprise data.
  • Automate data ingestion, storage, monitoring, alerting, and data quality validation workflows.
  • Tune Spark jobs for performance, low latency, high throughput, and cloud cost efficiency.
  • Develop CI/CD pipelines and Infrastructure-as-Code deployments using Terraform and Databricks CLI.
  • Build feature engineering pipelines and support model training, tracking, versioning, and deployment with MLflow.
  • Operationalize secure, scalable AI/ML models for client applications.
  • Enforce data security, access controls, compliance, lineage, metadata management, and documentation standards.
  • Collaborate with architects, data scientists, engineers, and client stakeholders to align solutions with business outcomes.

Requirements

  • 3+ years of hands-on data engineering experience focused on big data processing and cloud-native architectures.
  • 2+ years of hands-on Databricks experience, including Apache Spark, Delta Lake, and MLflow.
  • Databricks Certified Data Engineer Associate certification or higher.
  • Proficiency in Python, SQL, and Spark-based frameworks.
  • Experience developing and optimizing large-scale ETL/ELT pipelines.
  • Strong understanding of Lakehouse architecture and cloud-agnostic data solutions.
  • Familiarity with CI/CD and Infrastructure-as-Code for Databricks, including Terraform and Databricks CLI.
  • Knowledge of data governance, security, and compliance practices, including GDPR, HIPAA, or SOC 2.
  • Experience in Agile, DevOps, and MLOps environments.
  • Additional Databricks certifications, streaming technologies, orchestration tools, AI/ML integration, or client-facing consulting experience preferred.

Benefits

  • Remote-first work environment for US-based employees.
  • Paid time off.
  • Medical, dental, and vision coverage for full-time W2 positions.
  • Company-sponsored 401(k) retirement plan.
  • Annual learning and development stipend for courses, conferences, and certifications.
  • Peer mentorship, coaching, and professional growth opportunities.
  • Exposure to AWS GenAI, data, and cloud technologies across diverse industries.
  • Company outings and volunteer opportunities.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Data Engineer

phData 251-1K IT Services

Data Engineer at phData, a remote-first global data and AI consultancy, designing and delivering secure, scalable cloud data solutions for enterprise clients.

Apache Airflow Apache Spark AWS Azure Cassandra Databricks dbt Elasticsearch GCP Hadoop HDFS Kafka Luigi Python REST API Scala Snowflake Solr SQL System Design
1 day, 3 hours ago

Forward Deployed Data Engineer

Abacus Insights 51-250 Insurance

Abacus Insights is hiring a Forward Deployed Data Engineer to work alongside health plan clients, building governed data foundations, pipelines, and integrations that turn fragmented payer data into analytics-ready products.

Apache Spark Databricks Snowflake SQL
1 day, 3 hours ago

Data Engineer

Azumo 51-250 Internet Software & Services

Azumo is hiring a fully remote Latin America-based Data Engineer to build and operate production ingestion, transformation, storage, warehouse, and retrieval infrastructure that supports client AI systems.

Apache Airflow Apache Spark AWS Azure CI/CD Dagster Databricks dbt Docker Flink GCP Git GitHub Actions HIPAA Jupyter Kafka Prefect Python Snowflake SQL Terraform
1 day, 3 hours ago

[Job-32078] Mid Level/ Senior Data Engineer (Engenheiro de Dados), Brasil

CI&T 5K-10K Internet Software & Services

A CI&T busca uma pessoa engenheira de dados para construir e manter o pipeline assíncrono que conecta canais de atendimento a sistemas de IA, garantindo integridade, durabilidade e governança dos dados em ambientes OCI.

Kafka PostgreSQL Python Redis SQL
2 days, 2 hours ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers