Scrape Kickstarter campaigns across the whole site — all 15 categories and 154 subcategories — with funding goal, amount pledged, percent funded, backer count, creator history, launch date and deadline. Pick any mix of categories, or crawl everything.
Scrapes the complete Kith product catalog from Shopify — prices, per-size availability, sale/markdown flags, and variant data for resale tracking and arbitrage.
Scrape classified listings from Kleinanzeigen.de (formerly eBay Kleinanzeigen). Extract listing details including title, price, description, location, seller info, images, and category-specific attributes.
Scrape Japanese intercity bus data from Kosokubus.com. Covers Willer Express, JR Bus (all 6 regional subsidiaries), Sakura Kotsu, Keio Bus, Odakyu City Bus, Meitetsu Bus, and 30+ operators. Returns seat class, amenities, overnight flag, fares in JPY, and availability.
Scrape Malaysian transit data: KTMB rail stations (ETS, Intercity, Komuter — ~150 stations with IDs and state groupings) and Easybook intercity bus routes (80+ routes with distance, duration, and city-pair metadata). Transit network graph for travel apps, MaaS platforms, and mapping services.
Extract overseas property listings and the estate agencies behind them from Kyero across Spain, Portugal, France, Italy, Andorra, Greece, Cyprus and Turkey.
Extract attorney profiles, contact details, practice areas, and bios directly from law firm websites. Provide a list of law firm URLs and get structured attorney data including name, title, email, phone, education, bar admissions, and headshot.
Normalizes the scattered public GitHub collections of leaked/published AI system prompts (ChatGPT, Claude, Cursor, Devin, v0, Perplexity, and more) into one deduplicated dataset with product, vendor, version, and leak-date fields. Passive: reads only already-public repositories.
Scrape classified ads from leboncoin.fr — title, price, location, images, attributes, and owner type. Pass any category or search URL.
Scrape auction property listings from leilaoimovel.com.br — Brazil's largest real estate auction portal. Extracts sale price, appraisal value, discount percentage, modality, closing date, bank, property type, address, and more. Filter by state, city, property type, and bank.
Scrapes official school profiles and carnival results from LIESA, the governing body for Rio de Janeiro's Grupo Especial samba schools. Returns school history, colors, carnavalesco, interprete, samba-enredo title, ranking position, and scoring from liesa.com.br.
Scrapes employer reviews, ratings, and salary data from en-hyouban.com (エン カイシャの評判) — Japan's #2 employer-review platform. Extracts JSON-LD EmployerAggregateRating, per-dimension breakdowns, crowd-sourced salary, and employee reviews by URL or sitemap.
Scrape LinkedIn Ad Library ads by keyword, company ID, or country: advertiser name, ad creative, and the full EU DSA disclosure block — total and per-country impressions, audience targeting, and the legal entity paying for each ad.
Scrape public LinkedIn Services Marketplace listings across 200+ service categories (accountants, developers, designers, lawyers, and more). Returns provider name, category, headline, description, skills, location, rating, and review count. Auth-less, no login needed.
Scrape US federal lobbying disclosure filings from the Senate LDA database. Extract registrants, clients, lobbyists, activities, issue areas, and government entities. Filter by year, period, issue, registrant, or client.
Swiss business directory data: company names, addresses, phone numbers, emails, ratings, and opening hours for the Swiss SMB market. Filter by industry vertical. Structured JSON output ready for lead lists or CRM import.
Validate any NANP phone number and get back the real carrier of record, ILEC, rate centre, and a wireless-vs-landline flag from public numbering-plan data — plus local time and a TCPA calling-window check, not a format parse.
Scrape commercial property listings from LoopNet -- the #1 US commercial real estate marketplace. Extract addresses, prices, square footage, broker contacts, and more. Filter by property type (office, retail, industrial), listing type (for-sale/for-lease), and location.
Scrape used heavy equipment listings from MachineryTrader.com across every major equipment category (excavators, skid steers, dozers, cranes, loaders, and more). Get pricing, full specs, dealer contact info, location, and photos for every for-sale listing.
Cross-ecosystem feed of confirmed malicious/typosquatted packages across npm, PyPI, crates.io, Go, Maven, NuGet, Packagist and RubyGems, sourced from the OSSF malicious-packages dataset (also feeds OSV.dev). Category taxonomy plus typosquat-target linkage for CI gating. Passive read only.
Scrapes the Mararun platform — the dominant Chinese marathon management SaaS. Returns event details for mararun-hosted Chinese marathons: name, date, city, registration windows, participant cap, organizer, and CAA/AIMS certification.
Scrape the complete Maskota.com.mx product catalog — Mexico's leading online pet retailer. Extracts product names, prices, brands, variants, tags, and category data for all pets across the Shopify-hosted store.
Chile's national procurement source, ChileCompra Mercado Público — tenders, awards, bidder counts and purchase orders in one feed. Includes winning supplier RUT, buyer unit RUT, awarded amounts, and convenio-marco purchase orders that never appear as a tender.
Scrape MSD Saude Animal Brazil's authorized distributor network and veterinary product catalog. Returns contact details for 500+ distributors across 27 BR states by species category, plus regulatory data for 200+ vet pharma products.