Smart Working

Smart Working

Smart Working is a company that specializes in software development outsourcing and staff augmentation. They offer nearshore software development services, outsourcing solutions, and staff augmentation with a focus on providing highly skilled Indian de...

Internet Software & Services

Description

  • Architect and build scalable data pipelines and infrastructure to support AI and product systems.
  • Design and maintain data ingestion, transformation, and storage architectures for operational and AI workloads.
  • Develop and manage batch and real-time data pipelines.
  • Build and optimize systems for vector search, retrieval, and machine learning data pipelines.
  • Ensure data reliability, security, and governance across the platform.
  • Collaborate with AI, backend, product, and leadership teams to support training, inference, and product features.
  • Implement monitoring, observability, and data quality frameworks.
  • Optimize the performance of large-scale datasets and query systems.
  • Contribute to technical architecture decisions and long-term data strategy.
  • Act as the founding data hire and help define the culture, standards, and hiring bar for the growing data function.

Requirements

  • 7+ years of professional experience, with most of that in dedicated data engineering roles.
  • Strong experience designing and building data pipelines and distributed data systems.
  • Experience with relational databases, with PostgreSQL preferred; MySQL or similar acceptable.
  • Experience working with NoSQL databases.
  • Experience with vector databases used in modern AI systems.
  • Strong programming experience in Python.
  • Demonstrated ability to make and justify architectural decisions.
  • Experience building scalable backend systems.
  • Experience designing data models and storage architectures.
  • Strong understanding of data processing performance and optimization.
  • Experience with Apache Spark, Apache Airflow, Kafka, and Elasticsearch or OpenSearch is highly desirable.
  • Experience with PostgreSQL, MongoDB, and vector databases such as Qdrant, Milvus, or pgvector is highly desirable.
  • Experience with Python data-processing libraries such as Pandas or Polars is highly desirable.
  • Experience working on AI or machine learning platforms (nice to have).
  • Familiarity with stream processing and event-driven architectures (nice to have).
  • Experience with cloud infrastructure such as GCP, AWS, or Azure (nice to have).
  • Experience working in high-growth startups or early-stage companies (nice to have).

Benefits

  • Remote-first work environment.
  • Full-time, long-term role.
  • Opportunity to join a genuine community focused on growth and well-being.
  • Work with outstanding global teams and products.
  • Exposure to a high-rated workplace recognized on Glassdoor.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer II

GoGuardian 251-1K Internet Software & Services

GoGuardian is hiring a Data Engineer II to develop and improve the data platform, analytics, and AI/ML ecosystem supporting product insights and machine learning across the company.

Apache Airflow Apache Spark AWS CI/CD Dagster Databricks dbt Machine Learning MLOps Pandas Prefect Python SQL Terraform
49 minutes ago

[Job-31133] Senior/Mid Level, Data Engineer, Brazil

CI&T 5K-10K Internet Software & Services

CI&T is hiring a Mid/Senior Data Engineer to build scalable, intelligent data solutions for large volumes of data in a cloud and AI-driven environment.

Apache Spark AWS Datadog Kafka Python Serverless
1 hour, 34 minutes ago

Senior Data Engineer (GCP • Python • Iceberg • Delta Lake • Kafka • Snowflake • Databricks)

Railroad19 51-250 Internet Software & Services

Railroad19, Inc is hiring a Senior Data Engineer to design and deliver GCP-based lakehouse data-sharing and ingestion solutions for BigQuery, Snowflake, and Databricks.

Apache Spark Databricks GCP Kafka Python Snowflake
1 hour, 34 minutes ago

[Job - 31120] Mid Level Data Developer, Brazil

CI&T 5K-10K Internet Software & Services

CI&T is hiring a Data Developer to support a consumer electronics and audio technology client by maintaining reliable data models, transformations, and pipelines for analytics and reporting.

CI/CD dbt Git Snowflake SQL Tableau
1 hour, 34 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers