Senior Python Data Scraping Engineer (Freelance)

1 day, 8 hours ago
Part-time
Senior
Software Development
Mindrift.ai: Be the “I” in AI

Mindrift.ai: Be the “I” in AI

Join 10,000+ experts earning $15-50/hr training AI models remotely. Flexible freelance work, weekly payments. No AI experience required. Apply in 5 minutes.

Internet Software & Services

Description

  • Own end-to-end data extraction workflows across complex websites to deliver complete, accurate structured datasets.
  • Use internal tools such as Apify and OpenRouter alongside custom workflows to accelerate data collection, validation, and execution.
  • Handle dynamic and interactive web sources, including JavaScript-rendered content and changing site behavior.
  • Enforce data quality standards through validation checks, cross-source consistency controls, and systematic verification before delivery.
  • Scale scraping operations for large datasets using batching or parallelization while maintaining stability against minor site changes.
  • Monitor failures and troubleshoot issues to keep scraping workflows reliable.
  • Adapt extraction approaches for complex page structures, archived pages, and inconsistent HTML.
  • Work independently while applying critical thinking and domain expertise to AI-assisted tasks.

Requirements

  • 5+ years of relevant experience in data engineering, web scraping, automation, or software development.
  • Bachelor’s or Master’s degree in Engineering, Applied Mathematics, Computer Science, or a related technical field is a plus.
  • Strong technical foundation with practical experience in scripting, automation, and AI-assisted workflows.
  • Strong Python web scraping experience with tools such as BeautifulSoup and Selenium, including dynamic content, APIs, proxies, JS, AJAX, and infinite scroll.
  • Proven ability to extract data from complex structures such as hierarchies, archived pages, and inconsistent HTML.
  • Solid background in data cleaning, normalization, validation, and delivery of structured datasets in CSV, JSON, or Google Sheets.
  • Experience handling anti-bot mechanisms and dynamic site structures at scale.
  • Experience with cloud infrastructure such as AWS and containerization with Docker in real workflows.
  • Hands-on experience with LLM frameworks such as LangChain, OpenRouter, or similar.
  • English proficiency at upper-intermediate (B2) level or above.
  • Strong attention to detail and commitment to data accuracy.
  • Self-directed work ethic with the ability to troubleshoot independently.
  • A GitHub link is a plus.

Benefits

  • Part-time remote freelance opportunity.
  • Estimated workload of around 10–20 hours per week during active project phases.
  • Compensation of up to $45 per hour equivalent, depending on level and pace of contribution.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Software Engineer III, Content AI

Peloton 5K-10K Hotels, Restaurants & Leisure

Peloton is hiring a Software Engineer III in Warsaw to build and operate scalable Content AI microservices that power search, voice, and subtitles across its fitness platforms.

Agile AWS Bottle CI/CD Datadog Docker DynamoDB Elasticsearch GitHub Actions GraphQL gRPC Java Jenkins Kotlin Kubernetes Microservices MySQL NLP PostgreSQL Python Redis REST API Spring Boot Terraform
7 hours, 21 minutes ago

Senior Data Engineer

Sand Technologies Internet Software & Services

Sand Technologies is hiring a Senior Data Engineer to build the Azure data platform for UK utilities, starting with an AI-enabled situational awareness system for a major water utility.

Apache Spark Azure Databricks Flink IoT Kafka Machine Learning MQTT Python SFTP SQL
7 hours, 36 minutes ago

Sr. Data Engineer

Bloomerang 251-1K Specialized Consumer Services

Bloomerang is hiring a Senior Data Engineer to build and harden the unified data foundation behind its nonprofit giving platform, powering analytics, reporting, and AI products.

Apache Spark AWS CI/CD CRM Databricks dbt GCP Honeycomb Kafka Machine Learning MLflow OpenTelemetry Python Scala Snowflake SQL
7 hours, 51 minutes ago

Data Engineering Manager

Redwood Materials 251-1K Industrial Conglomerates

Redwood Materials is hiring a Data Engineering Manager to lead its data engineering and analytics team, own the data platform roadmap, and support the company’s battery supply chain operations through reliable data systems.

AWS AWS CDK Kubernetes Python SQL Trino
7 hours, 51 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers