Job Closed

This listing is no longer active.

Mindrift logo
Mindrift

Apply → Pass qualification(s) → Join a project → Complete tasks → Get paid. Project time expectations: Tasks are estimated to require around 10–20 hours per week during active phases, based on project requirements; This is an estimate, not a guaranteed workload, and applies only while the project is active. Note: Rates vary based on expertise, skills assessment, location, project needs, and other factors. Higher rates may be offered to highly specialized experts. Lower rates may apply during onboarding or non-core project phases. Payment details are shared per project.

Senior Data Scraping Engineer

Location

Worldwide

Posted

81 days ago

Salary

$37 / hour

Seniority

Senior

Job Description

Senior Data Scraping Engineer

Mindrift

Role Description Mindrift is looking for highly skilled Python Data Scraping Engineers to join the Tendem project and drive specialized data scraping workflows within our hybrid AI + human system. In this role, as an AI Pilot, you’ll collaborate with Tendem Agents that handle repetitive tasks, while you provide critical thinking, domain expertise, and quality control to deliver accurate and actionable results. This part-time remote opportunity is ideal for technical professionals with hands-on experience in web scraping, data extraction, and processing. Key Responsibilities - Own end-to-end data extraction workflows across complex websites, ensuring complete coverage, accuracy, and reliable delivery of structured datasets. - Leverage internal tools (Apify, OpenRouter) alongside custom workflows to accelerate data collection, validation, and task execution while meeting defined requirements. - Ensure reliable extraction from dynamic and interactive web sources, adapting approaches as needed to handle JavaScript-rendered content and changing site behavior. - Enforce data quality standards through validation checks, cross-source consistency controls, adherence to formatting specifications, and systematic verification prior to delivery. - Scale scraping operations for large datasets using efficient batching or parallelization, monitor failures, and maintain stability against minor site structure changes. Qualifications - At least 3 years of relevant experience in data engineering, web scraping, automation, or software development (required). - Bachelor's or Master’s Degree in Engineering, Applied Mathematics, Computer Science, or related technical fields is a plus. - Strong experience in Python web scraping (BeautifulSoup, Selenium or similar), including dynamic content (JS, AJAX, infinite scroll) and APIs via proxies. - Proven ability to extract data from complex structures (hierarchies, archived pages, inconsistent HTML). - Solid background in data cleaning, normalization, and validation, delivering structured datasets (CSV, JSON, Google Sheets). - Hands-on experience with LLMs and AI frameworks to enhance automation and problem-solving. - Strong attention to detail and commitment to data accuracy. - Self-directed work ethic with ability to troubleshoot independently. - A link to GitHub is a plus. - English proficiency: Upper-intermediate (B2) or above (required). Compensation On this project, contributors can earn up to $37 per hour equivalent, depending on their level and pace of contribution. Compensation varies across projects depending on scope, complexity, and required expertise. Please note that other projects on the platform may offer different earning levels based on their requirements. Benefits - Work fully remote on your own schedule with just a laptop and stable internet connection. - Gain hands-on experience in a unique hybrid environment where human expertise and AI agents collaborate seamlessly — a distinctive skill set in a rapidly growing field. - Participate in performance-based bonus programs that reward high-quality work and consistent delivery.

Related Categories

Related Job Pages

More Data Engineer Jobs

Full TimeRemoteTeam 1,001-5,000Since 1985H1B Sponsor

• Design and implement frameworks and features for the next generation of the Sophos Central platform, powered by the Taegis Data Platform • Build, maintain, and optimize large‑scale data pipelines to improve performance, reliability, and scalability • Participate in architecture, planning, and design discussions with engineering leaders, architects, and team leads • Collaborate closely with operations and site reliability teams to ensure solutions are production‑ready, observable, and supportable • Contribute to the ongoing evolution of platform standards around security, resiliency, and operational excellence

India
Job Closed
Zensar logo

Sr. Data Engineer

Zensar

At Zensar, we’re “experience-led everything”. We are committed to conceptualizing, designing, engineering, marketing, and managing digital solutions and experiences for over 130 leading enterprises. We are a company driven by a bold purpose: Together, we shape experiences for better futures. Whether for our clients, our people, or the world around us, this belief powers everything we do. At the heart of our culture is ONE with Client - a set of four core values that reflect who we are and how we work: One Zensar, Nurturing, Empowering, and Client Focus. Part of the $4.8 billion RPG Group, we’re a community of 10,000+ innovators across 30+ global locations, including Milpitas, Seattle, Princeton, Cape Town, London, Zurich, Singapore, and Mexico City. We believe the best work happens when individuality is celebrated, growth is encouraged, and well-being is prioritized. We are an equal employment opportunity (EEO) and affirmative action employer, committed to creating an inclusive workplace. All qualified applicants will be considered without regard to race, creed, color, ancestry, religion, sex, national origin, citizenship, age, sexual orientation, gender identity, disability, marital status, family medical leave status, or protected veteran status.

Data Engineer81 days ago
Full TimeRemoteTeam 10,001

Role Description - Python – strong hands-on experience in enterprise-scale development - Apache Spark / PySpark – building large-scale distributed data pipelines - Big Data concepts – partitioning, shuffling, performance tuning - SQL – complex queries, joins, and optimization - Cloud exposure on GCP, including: - BigQuery - Cloud Storage (GCS) - Dataflow (preferred) - Pub/Sub (basic understanding) - Experience working with large datasets (TB scale) Company Description At Zensar, we’re “experience-led everything”. We are committed to conceptualizing, designing, engineering, marketing, and managing digital solutions and experiences for over 130 leading enterprises. We are a company driven by a bold purpose: Together, we shape experiences for better futures. Whether for our clients, our people, or the world around us, this belief powers everything we do. - At the heart of our culture is ONE with Client - a set of four core values that reflect who we are and how we work: One Zensar, Nurturing, Empowering, and Client Focus. - Part of the $4.8 billion RPG Group, we’re a community of 10,000+ innovators across 30+ global locations, including Milpitas, Seattle, Princeton, Cape Town, London, Zurich, Singapore, and Mexico City. - Explore Life at Zensar and join us to Grow. Own. Achieve. Learn. to be the best version of yourself. - We believe the best work happens when individuality is celebrated, growth is encouraged, and well-being is prioritized. - We are an equal employment opportunity (EEO) and affirmative action employer, committed to creating an inclusive workplace. - All qualified applicants will be considered without regard to race, creed, color, ancestry, religion, sex, national origin, citizenship, age, sexual orientation, gender identity, disability, marital status, family medical leave status, or protected veteran status.

India
Job Closed
Full TimeRemoteTeam 201-500Since 2003H1B No Sponsor

• Own data quality: identify, analyze, and fix data issues • Translate business processes into reliable data structures • Build and optimize ETL pipelines • Deliver clean and structured data for analytics and BI • Handle AdHoc analytical requests under tight deadlines • Maintain and improve Power BI reporting • Automate repetitive tasks • Provide analytical support • Train users on BI tools

United States
Job Closed
Full TimeRemoteTeam 1,001-5,000H1B No Sponsor

• Build and maintain robust data pipelines processing large volumes of data • Analysis of large data sets using tools such as Python & SQL • Update and optimize our data platform for speed, scalability and cost • Coordinate with different functional teams to understand and meet their data needs • Develop processes and tools to monitor and analyze model performance and data accuracy • Solve general data-related problems • Setting up new pipelines for the full stream/enrichment/curation process • Upkeep of source code locations • Investigating and utilising ML & AI to improve the cloud offering • Development of junior staff members

Spain
Job Closed