Scrape listings from Jmty (ジモティー), Japan's dominant hyper-local classifieds and flea-market platform. Extract titles, prices, locations, categories, descriptions, images, and seller info by prefecture and category.
Scrapes Indonesian public transit: KAI (PT Kereta Api Indonesia) station catalog (217 stations across Java, Sumatra, Sulawesi) with codes and city data, plus TransJakarta BRT corridor listing. For travel apps, logistics planners, and Indonesian transit integrations.
Scrapes Keep (gotokeep.com) — China's largest fitness platform with 200M+ users. Extracts yoga, meditation, qigong, HIIT and other fitness courses including title, difficulty level, participant count, equipment required, cover image, and category grouping.
Extract vehicle valuations from Kelley Blue Book. Get fair market price, MSRP, trade-in value, consumer reviews, specs, and trim details for any car by year, make, and model. Built for dealers, researchers, and automotive market analysis.
Scrape Kew POWO — the global authoritative plant taxonomy backbone (1.4M names, WCVP). Returns full classification, accepted/synonym status, native distribution, lifeform, synonyms, and child taxa. Search by name, family, or genus.
Scrapes the complete Kith product catalog from Shopify — prices, per-size availability, sale/markdown flags, and variant data for resale tracking and arbitrage.
Scrape Japanese intercity bus data from Kosokubus.com. Covers Willer Express, JR Bus (all 6 regional subsidiaries), Sakura Kotsu, Keio Bus, Odakyu City Bus, Meitetsu Bus, and 30+ operators. Returns seat class, amenities, overnight flag, fares in JPY, and availability.
Scrape Malaysian transit data: KTMB rail stations (ETS, Intercity, Komuter — ~150 stations with IDs and state groupings) and Easybook intercity bus routes (80+ routes with distance, duration, and city-pair metadata). Transit network graph for travel apps, MaaS platforms, and mapping services.
Extract attorney profiles, contact details, practice areas, and bios directly from law firm websites. Provide a list of law firm URLs and get structured attorney data including name, title, email, phone, education, bar admissions, and headshot.
Scrapes official school profiles and carnival results from LIESA, the governing body for Rio de Janeiro's Grupo Especial samba schools. Returns school history, colors, carnavalesco, interprete, samba-enredo title, ranking position, and scoring from liesa.com.br.
Scrape US federal lobbying disclosure filings from the Senate LDA database. Extract registrants, clients, lobbyists, activities, issue areas, and government entities. Filter by year, period, issue, registrant, or client.
Scrape commercial property listings from LoopNet -- the #1 US commercial real estate marketplace. Extract addresses, prices, square footage, broker contacts, and more. Filter by property type (office, retail, industrial), listing type (for-sale/for-lease), and location.
Scrape real-time China box office data from Maoyan Piaofang. Extracts daily revenue, market share, screening counts, seat occupancy, and cumulative totals for all currently-screening films.
Extract finisher results from MarathonGuide — the largest US road-race results database covering thousands of marathons and road races since the 1990s. Search by race name or year to retrieve finish times, placements, age-group rankings, and runner details.
Pull structured source-credibility records from Media Bias/Fact Check (MBFC) -- the largest media-source reliability database (~7,000+ profiles). Returns bias rating, factual-reporting tier, MBFC credibility rating, country press-freedom, media type, traffic, and full History/Funding/Analysis prose.
Scrape MSD Saude Animal Brazil's authorized distributor network and veterinary product catalog. Returns contact details for 500+ distributors across 27 BR states by species category, plus regulatory data for 200+ vet pharma products.
Scrapes mid-career job postings from Mynavi Tenshoku (tenshoku.mynavi.jp), Japan's top-3 転職 board. Returns structured records with salary ranges, work hours, holidays/vacation, and required skills — useful for JP labor-market analysis, comp benchmarking, and recruiter research.
Scrapes the Mystic Stamp Company US catalog (mysticstamp.com). Returns 60k+ US postage stamp listings keyed to Scott catalog numbers — the standard US philatelic reference. Each record includes Scott number, title, issue year, denomination, price, stock status, and image URL.
Scrape live UK rail data from National Rail: departure and arrival boards by station, real-time delays, cancellations, platforms, and operator info. Covers GWR, LNER, Avanti, Southern, Thameslink, ScotRail, TfW, Elizabeth line, CrossCountry, Northern, and more.
Extract Netflix's official weekly Top 10 rankings across 90+ countries and globally. Pulls full historical data from Netflix Tudum's TSV files: all-weeks-global (hours viewed, views, runtime) and all-weeks-countries (since 2021). One record per ranked title per week.
Scrape NIH-funded research projects from the official RePORTER v2 API. Extract PI names, award amounts, activity codes (R01, R21, K99), study sections, dates, and active/terminated status. Optionally pull linked publications (PMIDs). Filter by keyword, fiscal year, PI, org, state, or institute.
Scrape free US people data from Nuwber — full name, age, phone numbers, current and past addresses, and relatives. Search by first/last name with optional state filter. Bypasses Cloudflare protection via a real-browser (Camoufox) fetch.
Extract sanctioned entities from the US Treasury OFAC SDN and Consolidated lists, plus optional multi-jurisdiction coverage (UK, EU, UN, Canada, Australia, Switzerland, Japan, Israel). Get names, aliases, addresses, programs, and vessel data. Filter by type, program, country, or name.
Resolve author names to their full Open Library bibliography: works list with first-publish year, edition count, subjects, and cover URLs. Returns bio, birth/death dates, external IDs (VIAF, Wikidata, ISNI), and alternate names. For literary databases, recommendation engines, and RAG pipelines.