Scrapes DxOMark lab-bench scores for cameras, lenses, and sensors. Extracts overall scores, sub-scores (Portrait/Landscape/Sports for sensors; Sharpness/Transmission/Distortion/Vignetting/CA for lenses), review dates, and product metadata. Walks the DxOMark sitemap to discover all products.
Scrapes scores, schedules, rosters, and athlete data from ESPN's undocumented public JSON APIs (site.api.espn.com + sports.core.api.espn.com). Covers NFL, NBA, MLB, NHL, soccer, and college sports — no auth, no HTML parsing, survives site redesigns.
Scrape tours, activities, and experiences from GetYourGuide.com by search query or direct URL.
Extracts title, price, rating, review count, images, highlights, duration, cancellation policy,
location, and category. Handles Cloudflare bot protection automatically.
Classified ad and second-hand listings from Gumtree UK, Australia and South Africa. Search by keyword, browse a category, or supply listing URLs. Returns title, price, location, description, images, seller name and type, attributes and date posted. For local pricing and resale sourcing.
Scrape job listings from Gupy — Brazil's #1 ATS used by 3,000+ enterprises including Itaú, Petrobras, and Magalu. Search by keyword, workplace type, or state. Returns job title, company, location, remote status, PCD flag, description, and apply URL.
Scrape for-sale listings from LIFULL HOMES (homes.co.jp). Covers resale condos (中古マンション), new condos, detached homes, and land. Extracts price in JPY and USD, floor area, layout, management fees, station access, zoning, photos, and broker info.
Scrape the HubSpot App Marketplace catalog. Returns app name, category, developer, rating, reviews, pricing model, HubSpot certification flag, and Hub compatibility (Marketing, Sales, Service, CMS, Operations). For agencies, app builders, ecosystem research, and partner analytics.
Scrape job listings from InfoJobs.net (Spain) — extract titles, companies, locations, salaries, contract types, experience requirements, skills, and full descriptions.
Scrapes the official Innpulsa Colombia convocatorias catalog (Colombia's flagship gov entrepreneurship-acceleration body). Returns all open and closed grant calls with full metadata: dates, eligibility, benefits, registration links, and document terms.
Scrapes the US insurance carrier directory from the NAIC Consumer Information Source — ~5,200 licensed carriers with NAIC code, carrier name, licensed states, lines of business, address, phone, and website. Covers PC, life, health, surplus-lines, and captive carriers.
Scrapes the italki language-teacher marketplace directory. Extracts teacher profiles including pricing, ratings, lesson counts, teaching languages, and spoken languages from the italki SSR pages. Supports filtering by language and limiting result count.
Scrape the full Johnny's Selected Seeds catalog — ~3,300 SKUs across vegetables, herbs, flowers, fruit, and cover crops. Extracts agronomic data: botanical name, days to maturity, life cycle, disease resistance codes, organic/hybrid status, pack-size pricing, and growing information.
Scrapes Jomashop's public Magento GraphQL API for luxury watch, jewelry, fragrance, and apparel listings — including grey-market discounted prices, MSRP anchors, and the discount delta. Supports full-category firehose or targeted product/category lookup.
Scrapes daily and weekly Spotify charts for 7 LATAM countries (BR, MX, AR, CL, CO, PE, UY) from kworb.net. Returns ranked tracks with stream counts, peak positions, days on chart, and Spotify track IDs.
Scrape LATAM startup funding rounds from LatamList and Contxto. Extracts company name, country, sector, stage, amount, lead investor, and participants from funding articles. Covers Brazil, Mexico, Colombia, Argentina, Chile, and more. Ideal for VCs and founders tracking LATAM deal flow.
Scrape public room metadata from LINE OpenChat — Japan's largest public chat platform, also popular in Taiwan and Thailand. Extracts room name, description, member count, hashtags, region, and more from the public OpenChat directory. No LINE account required. Supports JP, TW, and TH markets.
Scrape product listings and supplier data from Made-in-China.com by keyword. Collect product names, prices, MOQ, supplier names, and profile links — ideal for importers, sourcing agents, and procurement analysts.
Scrape public rooms and messages from any Matrix homeserver (matrix.org, Element, or self-hosted). Discover public rooms by keyword or scrape message history from specific rooms using a Matrix access token.
Scrape Metacritic for Metascores, user scores, critic review counts, genres, directors and full metadata for movies and TV shows. Covers the full catalog via sitemaps — ~30k movies and ~12k TV titles with a unified schema.
Bulk scraper for Mexico's DENUE business registry — 5.5M+ establishments with geo coordinates, SCIAN industry codes, employee ranges, contact info, and addresses. Filter by state, activity keyword, and entity type. Requires a free INEGI API token (email registration at inegi.org.mx).
Scrape job listings from Monster.com by keyword and location. Extract job title, company, salary, location, employment type, date posted, and more. Supports US and international markets.
Extract US mine records from MSHA Open Government Data CSVs. Covers coal, metal, and nonmetal mines: mine ID, operator, controller, state/county, geocoordinates, mine type, commodity, status, employees, quarterly production tons, violations, and accidents.
Scrape Indian real estate listings from NoBroker.in by city/locality search URL. Returns price, BHK type, area, address, amenities, owner details, and property URL for rent and sale listings.
Scrape NYC building violations, permits, and property data from 7 open city datasets (DOB violations, ECB violations, HPD housing violations, permits, job filings, PLUTO property profile, LL84 energy benchmarking) joined on BIN/BBL into a single building-risk profile per address.