Smart Working

Smart Working

Smart Working is a company that specializes in software development outsourcing and staff augmentation. They offer nearshore software development services, outsourcing solutions, and staff augmentation with a focus on providing highly skilled Indian de...

Internet Software & Services

Description

  • Architect and build scalable data pipelines and infrastructure for AI and product systems.
  • Design and maintain data ingestion, transformation, and storage architectures for operational and AI workloads.
  • Develop and manage batch and real-time data pipelines.
  • Build and optimize systems for vector search, retrieval, and machine learning data workflows.
  • Ensure data reliability, security, and governance across the platform.
  • Collaborate with AI, backend, product, and leadership teams to support training, inference, and product features.
  • Implement monitoring, observability, and data quality frameworks.
  • Optimize performance for large-scale datasets and query systems.
  • Contribute to technical architecture decisions and long-term data strategy.
  • Define the data function, establish standards, and help build the team as the company scales.

Requirements

  • 7+ years of professional experience, with most of it in dedicated data engineering roles.
  • Strong experience designing and building data pipelines and distributed data systems.
  • Experience with relational databases, with PostgreSQL preferred but MySQL or similar acceptable.
  • Experience working with NoSQL databases.
  • Experience with vector databases used in modern AI systems.
  • Strong programming experience in Python.
  • Demonstrated ability to make and justify architectural decisions.
  • Experience building scalable backend systems.
  • Experience designing data models and storage architectures.
  • Strong understanding of data processing performance and optimization.
  • Experience with Apache Spark, Apache Airflow, Kafka, and Elasticsearch or OpenSearch is highly desirable.
  • Experience with PostgreSQL, MongoDB, and vector databases such as Qdrant, Milvus, or pgvector is highly desirable.
  • Experience with Python data-processing libraries such as Pandas or Polars is highly desirable.
  • Experience working on AI or machine learning platforms is preferred.
  • Familiarity with stream processing and event-driven architectures is preferred.
  • Experience with cloud infrastructure such as GCP, AWS, or Azure is preferred.
  • Experience working in high-growth startups or early-stage companies is preferred.

Benefits

  • Remote-first work environment.
  • Opportunity to join a genuine community focused on growth and well-being.
  • Full-time, long-term role with a global team.
  • Work with a company that values personal and professional growth.
  • Join one of the highest-rated workplaces on Glassdoor.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Data Engineer IV (6329)

Dan.com - a GoDaddy brand Internet Software & Services

itD is seeking a remote Data Engineer IV for a 12-month engagement to scale data infrastructure, enable AI-driven analytics, and deliver customer experience insights for its Post Sales Organization.

Apache Airflow Apache Spark CI/CD Feature Engineering Generative AI Git Google Analytics Kafka Python Snowflake SQL Tableau
9 hours, 24 minutes ago

(1573) Data Engineer (PostgreSQL & Databricks)

Nortal 251-1K Internet Software & Services

Nortal is building a LATAM remote talent network for Data Engineers to design, develop, maintain, and support scalable data pipelines and customer-facing applications using Databricks, Postgres, and SQL Server.

Agile AWS CI/CD Databricks DevSecOps Microservices PostgreSQL Scrum Serverless SQL Server
9 hours, 39 minutes ago

Data Platform Engineer

Figma 1K-5K Internet Software & Services

Figma is hiring a Data Platform Engineer to build foundational data and infrastructure systems that support analytics, machine learning, and AI-driven product experiences.

CI/CD Machine Learning Python
10 hours, 24 minutes ago

Marketing Data Engineer - Ad Tech (US hours)

VirtuHire 11-50 Staffing / Recruiting

The Marketing Data Engineer will own the production data pipelines and data-quality foundation that support dashboards, pacing outputs, and recurring reporting for a U.S.-based client.

Python Snowflake SQL
10 hours, 54 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers