Scrape Metacritic for Metascores, user scores, critic review counts, genres, directors and full metadata for movies and TV shows. Covers the full catalog via sitemaps — ~30k movies and ~12k TV titles with a unified schema.
Est. monthly creator revenue
- Low
- $0.23
- Central
- $0.97
- High
- $9.07
Extract Netflix's official weekly Top 10 rankings across 90+ countries and globally. Pulls full historical data from Netflix Tudum's TSV files: all-weeks-global (hours viewed, views, runtime) and all-weeks-countries (since 2021). One record per ranked title per week.
Scrape NYC building violations, permits, and property data from 7 open city datasets (DOB violations, ECB violations, HPD housing violations, permits, job filings, PLUTO property profile, LL84 energy benchmarking) joined on BIN/BBL into a single building-risk profile per address.
Scrapes OFA breed-level disease-prevalence statistics for hip dysplasia, elbow dysplasia, cardiac disease, patellar luxation, thyroid, eyes, and DNA tests. Returns per-breed evaluation counts and percentage breakdowns — used by pet-insurance actuaries and AKC breeder compliance auditors.
Scrape ebook and audiobook availability from any public library's OverDrive/Libby catalog. Provide library slugs and search queries — get real-time copy counts, holds, wait times, and full metadata. The digital-circulation signal publishers and authors can't source elsewhere.
Scrape the Preply tutor directory for language and subject tutors. Extracts name, subjects, native language, country of birth, hourly rate (USD), trial lesson price, rating, review count, lesson count, languages spoken, profile URL, and badges.
Scrape speedrun.com leaderboard rankings for any game and category. Resolves game names to IDs, enumerates categories automatically, and returns flattened run rows with player names, times, ranks, WR flags, platform, region, and video links via the official v1 REST API.
Generate realistic e-commerce test data with interconnected products, customers, orders, and reviews. Features referential integrity, realistic distributions, temporal coherence, industry presets, and deterministic seed mode.
Extract Thailand transit data from three sources: SRT station catalog (~750 stations, bilingual Thai/English), BTS Skytrain stations with first/last train times (~119 stations), and ARL Red Line departure timetables. Outputs line assignments and service direction data.
Scrape full restaurant menus from UberEats. Extract restaurant info, all menu sections, items with prices, descriptions, and images.
Scrapes Ukraine ProZorro public-procurement contracts in reconstruction CPV / DK021 categories (construction, engineering, demolition, capital repair). Emits buyer, supplier, EDRPOU codes, contract value, dates, item descriptions. No auth.
Stream of Executive Orders, Proclamations, Memoranda, Determinations, and Notices from the Federal Register with markdown full text, citations, and signing dates. Filter by date, president, action type, or keyword.
Scrapes the Lady Bird Johnson Wildflower Center NPIN — the canonical US native-plant database with ~9,000 species. Returns scientific name, common names, USDA symbol, growing conditions, bloom data, wildlife value, pollinator and butterfly larval-host data, and commercial availability.
Scrape Philippines transit: inter-island ferries (2GO, Montenegro, Oceanjet, FastCat, Starlite) with cabin-class breakdown and RORO vehicle fares in PHP, PNR/MRT-3/LRT-1/LRT-2 rail, PITX provincial bus, and Cebu Pacific domestic flights.
Pull voluntary carbon offset and CDR projects from three top registries: American Carbon Registry (ACR), Climate Action Reserve (CAR), and Puro.earth (carbon dioxide removal). Returns project type, methodology, developer, country, crediting period, credits issued, and status. No auth required.
Scrapes ai-bot.cn — the definitive curated China AI tool directory — into a clean dataset of 1,000+ tools with names, categories, descriptions, logos, and external links. Covers domestic (China) and overseas AI tools with origin classification.
Short-term and vacation rental listing data, by location. Returns nightly price, property and room type, amenities, host and superhost status, rating, review count, guest capacity, photos and coordinates. Optional dates and guest count. For rental market research, pricing benchmarks and host lists.
Track Amazon Best Sellers rankings across any category. Collects product name, rank, ASIN, price, rating, review count, and image URL from any bestsellers category page.
Scrape US intercity bus schedules and fares: Megabus and FlixBus (Greyhound/FlixBus network) with per-carrier trip IDs, origin/destination cities, departure/arrival times, fares, seat availability, amenities, and booking links. Covers 734+ city pairs and 500+ routes across the continental US.
Extract German job listings from Arbeitsagentur.de (Bundesagentur für Arbeit), Germany's federal employment agency. Search by keyword, location, or browse all open positions. Returns job title, employer, location, contract type, posted date, and job URL.
Scrapes the ASPCA's canonical toxic and non-toxic plants database — the gold-standard pet-safety reference. Returns tri-species toxicity data (dog/cat/horse), scientific names, families, toxic principles, clinical signs, and images for ~1,024 plants.
Scrapes California State Bar disciplinary actions — disbarments, suspensions, probations, censures. One row per action with attorney name, bar number, sanction type, effective date, violations, and profile link.
Scrape Audubon's "Plants for Birds" database by US ZIP code — returns native plants ranked for bird habitat value with bird species attracted, plant type, and wildlife resources provided. Unique zip → native-plant → bird-ecology join not available elsewhere as a structured data feed.
Scrape Behance project metadata including title, owner, tags, engagement stats, and cover image URL. Search by keyword — illustrations, paintings, branding, and more.