Extract attorney profiles, contact details, practice areas, and bios directly from law firm websites. Provide a list of law firm URLs and get structured attorney data including name, title, email, phone, education, bar admissions, and headshot.
Scrapes official school profiles and carnival results from LIESA, the governing body for Rio de Janeiro's Grupo Especial samba schools. Returns school history, colors, carnavalesco, interprete, samba-enredo title, ranking position, and scoring from liesa.com.br.
Scrape commercial property listings from LoopNet -- the #1 US commercial real estate marketplace. Extract addresses, prices, square footage, broker contacts, and more. Filter by property type (office, retail, industrial), listing type (for-sale/for-lease), and location.
Scrape real-time China box office data from Maoyan Piaofang. Extracts daily revenue, market share, screening counts, seat occupancy, and cumulative totals for all currently-screening films.
Extract finisher results from MarathonGuide — the largest US road-race results database covering thousands of marathons and road races since the 1990s. Search by race name or year to retrieve finish times, placements, age-group rankings, and runner details.
Pull structured source-credibility records from Media Bias/Fact Check (MBFC) -- the largest media-source reliability database (~7,000+ profiles). Returns bias rating, factual-reporting tier, MBFC credibility rating, country press-freedom, media type, traffic, and full History/Funding/Analysis prose.
Scrape MSD Saude Animal Brazil's authorized distributor network and veterinary product catalog. Returns contact details for 500+ distributors across 27 BR states by species category, plus regulatory data for 200+ vet pharma products.
Scrape live UK rail data from National Rail: departure and arrival boards by station, real-time delays, cancellations, platforms, and operator info. Covers GWR, LNER, Avanti, Southern, Thameslink, ScotRail, TfW, Elizabeth line, CrossCountry, Northern, and more.
Extract Netflix's official weekly Top 10 rankings across 90+ countries and globally. Pulls full historical data from Netflix Tudum's TSV files: all-weeks-global (hours viewed, views, runtime) and all-weeks-countries (since 2021). One record per ranked title per week.
Scrape NIH-funded research projects from the official RePORTER v2 API. Extract PI names, award amounts, activity codes (R01, R21, K99), study sections, dates, and active/terminated status. Optionally pull linked publications (PMIDs). Filter by keyword, fiscal year, PI, org, state, or institute.
Extract sanctioned entities from the US Treasury OFAC SDN and Consolidated lists, plus optional multi-jurisdiction coverage (UK, EU, UN, Canada, Australia, Switzerland, Japan, Israel). Get names, aliases, addresses, programs, and vessel data. Filter by type, program, country, or name.
Resolve author names to their full Open Library bibliography: works list with first-publish year, edition count, subjects, and cover URLs. Returns bio, birth/death dates, external IDs (VIAF, Wikidata, ISNI), and alternate names. For literary databases, recommendation engines, and RAG pipelines.
Scrape the Preply tutor directory for language and subject tutors. Extracts name, subjects, native language, country of birth, hourly rate (USD), trial lesson price, rating, review count, lesson count, languages spoken, profile URL, and badges.
Extract US real estate leads from Propwire with 11M+ properties including owner info, equity, MLS data, lead-type flags, and tax records.
Scrape live campsite availability from Recreation.gov's public API for any federal campground. Supply campground IDs and date ranges to get per-day open/reserved/closed status for every campsite — ideal for cancellation sniping, trip-planning feeds, and monitoring workflows.
Scrape yoga, meditation, ayahuasca, and wellness retreat listings from retreat.guru across 12 Latin American countries. Extracts event dates, pricing, ratings, location, and teacher info via inline JSON-LD on each event page.
Scrape live listings from Reverb's marketplace API. Search by keyword, category, or condition — get structured data on price, seller, photos, shipping, and more from the music gear marketplace.
Scrape heavy equipment auction listings from Ritchie Bros (rbauction.com). Extract lot details including manufacturer, model, year, serial number, usage hours, location, auction dates, bid prices, and condition flags across all equipment categories.
Scrape original artwork listings from Saatchi Art. Extract title, artist, price, style, medium, dimensions, and availability from 2M+ artworks. Filter by category, style, subject, or supply your own listing URLs.
Parse every position from SEC Form 13F-HR filings into structured rows: CUSIP, issuer, shares, USD value, voting authority, put/call. Query by filer, ticker, CUSIP, or reporting quarter.
Scrape Snapchat Spotlight videos by hashtag. Extracts video metadata, engagement stats (views, likes, shares, comments), creator info, and video URLs from public Spotlight feeds.
Scrape speedrun.com leaderboard rankings for any game and category. Resolves game names to IDs, enumerates categories automatically, and returns flattened run rows with player names, times, ranks, WR flags, platform, region, and video links via the official v1 REST API.
Scrapes stamp listings from the Stanley Gibbons online shop — the canonical British and Commonwealth philatelic catalogue authority. Extracts SG catalogue numbers, conditions, issue years, prices, and availability across all stamp collections.
Extract Thailand transit data from three sources: SRT station catalog (~750 stations, bilingual Thai/English), BTS Skytrain stations with first/last train times (~119 stations), and ARL Red Line departure timetables. Outputs fares in THB, line assignments, and service direction data.