Sonatype

Sonatype

Sonatype provides secure software development solutions by leveraging open source and artificial intelligence, ensuring that organizations can build applications quickly and safely through automated governance, policy enforcement, and comprehensive mon...

Internet Software & Services
51-250
Founded 2008
$155M raised

Description

  • Build and maintain reliable data pipelines and ETL/ELT workflows.
  • Develop and optimize data models for analytics and internal tools.
  • Collaborate with team members to deliver clean, trusted datasets.
  • Support Spark and AWS data platform services, including S3, SNS, SQS, ECS/Fargate, and EMR.
  • Monitor pipeline quality, performance, and reliability.
  • Document systems and contribute to test coverage and CI/CD processes.
  • Help shape the data lakehouse architecture and platform roadmap.

Requirements

  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • 2–4 years of experience in data engineering or a backend data-related role.
  • Strong skills in Java, Scala, or another backend programming language.
  • Python experience with PySpark or pandas.
  • Experience with SQL and distributed data systems such as Spark, Kafka, or SQS.
  • Familiarity with NoSQL databases such as Cassandra or HBase.
  • Understanding of data modeling for analytics and reporting.
  • Strong English communication skills, including the ability to explain work to non-engineers.
  • Self-directed, methodical problem-solving approach and openness to code review feedback.
  • Experience with AI-assisted engineering tools; dbt, Databricks, real-time pipelines, Terraform, CloudFormation, data governance, ML pipelines, or compliance standards are preferred.

Benefits

  • Flexible working practices.
  • Parental leave.
  • Diversity and inclusion working groups.
  • Cloud-first, open-source-friendly work environment.
  • Opportunities to work on data supporting software security outcomes.
  • Team culture emphasizing clarity, learning, and autonomy.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Senior Data Engineer (Databricks)

Livefront 11-50 Internet Software & Services

Livefront is hiring a Senior Databricks Data Engineer to build production-grade data and AI foundations that power clients’ digital products, analytics, and intelligent experiences.

Apache Spark AWS Azure CI/CD Databricks dbt GCP Git Go Java Kafka MLflow MLOps Python Scala SQL
9 hours, 24 minutes ago

[Job-31365] Data Developer (Azure/Databrics) Sênior , Brazil

CI&T 5K-10K Internet Software & Services

A CI&T busca uma pessoa Engenheira de Dados para atuar na transformação de cadastro de clientes em um contexto financeiro, desenvolvendo soluções de dados e apoiando a implantação de modelos de IA em nuvem.

Apache Spark Azure CI/CD Databricks GitHub MLflow Python SQL
9 hours, 54 minutes ago

Senior Consulting Engineer

MongoDB 1K-5K Internet Software & Services

MongoDB is hiring a remote Data Migration Consultant in Germany to lead enterprise migrations from relational and NoSQL platforms to MongoDB while developing scalable delivery frameworks and tooling.

Apache Spark AWS Azure C# GCP Go Java Kafka MongoDB MySQL Node.js Oracle PostgreSQL Python SQL Server
9 hours, 54 minutes ago

[Job-31470] Data Developer (Azure/Databrics) Sênior , Brazil

CI&T 5K-10K Internet Software & Services

A CI&T busca uma pessoa Engenheira de Dados para atuar em um projeto do setor financeiro, desenvolvendo e mantendo soluções de dados e IA em ambiente Azure e Databricks.

Apache Spark Azure CI/CD Databricks GitHub MLflow Python SQL
10 hours, 24 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers