Data Type
Task Types
Subject Matter / Industry
Language
Service Model
Share link
Dataset Description (5–8 words): Structured data from complex websites
Data Type (select one): Text
Subject Matter/Industry (5–8 words): Web data extraction and automation
Pre-labeled Data (Yes/No): No
Labeling Software: Other
Label Types (select at least 1):
Labeling Overview: You should have hands-on experience with Python-based web scraping and data extraction from complex sites, including dynamic/JavaScript-rendered pages. You’ll be comfortable troubleshooting scraping failures, validating outputs, and delivering clean structured data. Upper-intermediate English (B2) or higher is required.
In this role, you’ll own end-to-end scraping workflows: extracting data across multi-level site structures, using a mix of internal tools (Apify, OpenRouter) and your own scripts/workflows. You’ll validate and normalize data, enforce formatting requirements, and deliver accurate structured datasets (e.g., CSV/JSON/Sheets). You’ll collaborate in a hybrid AI + human setup where AI agents handle repetitive steps and you provide quality control and critical thinking.
Required Locations: Global - Any Location
Required English Level: Fluent
Other Qualifications & Requirements (5–10 bullets):