Pull country-level electricity data from Ember — generation TWh by fuel, capacity GW, CO2 intensity gCO2/kWh, demand across ~250 countries. Yearly (1985+) and monthly (2010+). The leading free substitute for IEA's paywalled Electricity Information dataset. No auth.
Scrape active business-for-sale listings from Empire Flippers marketplace. Extracts listing number, niche, monetization, asking price, monthly net profit, monthly revenue, multiple, age, and status.
Crawl toxic chemical release data from the EPA TRI via the Envirofacts API. Extract facility details, chemical names, release quantities by media (air, water, land), coordinates, and carcinogen flags. Filter by state, chemical, year, and facility.
Scrape the EU Transparency Register: organisations, financial disclosure, clients, accredited lobbyists, and Commissioner meetings. Ideal for public-affairs intelligence, ESG screening, political-risk research, and compliance due diligence.
Scrape company listings from Europages, Europe's largest B2B supplier directory. Search by product, service, country, and company type. Extract company name, address, phone, website, products, certifications, and employee range.
Scrape expert witness profiles from Experts.com — the largest open expert witness directory. Extracts contact info, specialty, credentials, categories, and biography for thousands of expert witnesses across all legal specialties.
Scrapes the FAA Office of Commercial Space Transportation (AST) website for commercial space launch license actions, program pages, and authorization announcements. Returns structured records from FAA press releases and stakeholder engagement pages covering licensed launch and reentry operations.
Scrape Failory's live startups directory — 14,000+ startups across 267+ country, city, and industry facet pages. Extracts startup name, website URL, industry, year founded, funding amount, funding round, and facet label. Ideal for lead generation, VC research, and competitive intelligence.
Scrapes FCC IBFS satellite service filings from fcc.report. Covers earth-station (SES) and space-station (SAT) applications with applicant, status, frequency bands, attached document URLs, and public notices. Supports date-range filtering for incremental runs.
Scrapes FishBase — 35,000+ fish species with taxonomy, ecology, size/weight, trophic level, IUCN status, game-fish flag, depth range, and common names.
Scrapes Flickr's camera-finder popularity dataset — photo upload counts, photographer user counts, and popularity scores for 600+ camera models across 50+ brands. Uses Flickr's internal camerafinder API to deliver the canonical real-world camera adoption signal.
Scrapes ~800 franchise brands from FranchiseDirect.com. Extracts franchise fee, minimum cash required, total investment range, territory availability, and brand description across all industry categories.
Extract book data from Goodreads: titles, authors, ratings, reviews, genres, ISBN, pages, format, publication date, awards, and more. Accepts book or author URLs as input.
Search and scrape US court opinions from Google Scholar's case law database. Filter by keywords, courts, and date ranges to collect case names, citations, courts, dates, snippets, "Cited by" counts, and links. Optionally fetch full opinion text from detail pages.
Grants.gov scraper for federal grant opportunities (posted, forecasted, closed, archived). Crawl titles, agencies, CFDA numbers, funding amounts, eligibility, close dates, and grantor contacts. Filter by agency, category, instrument, eligibility, CFDA, keyword, or posted-date window.
Scrapes race info, packages, registration details, and results history from the official Great Wall Marathon site (great-wall-marathon.com). Returns structured records for the current race edition, package pricing, and historical result editions.
Extracts the worldwide directory of Herman Miller authorized dealers, retailers, showrooms, and resellers from the public dealer locator API. Returns name, address, contacts, certification level, product lines, and business categories for every location globally.
Scrapes HobbyLink Japan (hlj.com), the largest English-language Japanese hobby shop, for Gunpla and anime figure catalog data: names, prices in JPY, release status, manufacturer, series, scale, JAN barcodes, and image URLs. Category backfill, new-release weekly, and pre-order modes.
Compare hotel prices across Booking.com, Expedia, Hotels.com, and Priceline from a single search. Collect nightly rates, guest ratings, star ratings, amenities, and photos for any destination. Perfect for travel market research and rate intelligence.
Scrape AI/ML model metadata from the HuggingFace Hub. Extract model names, task types, download counts, likes, libraries, authors, tags, licenses, model sizes, and model card excerpts. Filter by task type, library, author, and search query.
Aggregates 6 major HVAC OEM dealer networks (Trane, Lennox, Carrier, Rheem, Goodman, American Standard) into one structured feed with manufacturer certification tier attached.
Search the iTunes/Apple Music catalog by keyword, artist, or album; look up tracks by Apple ID; and fetch Top Charts rankings for any country. Metadata and 30-second preview URLs only — no audio downloads.
Scrape Kew POWO — the global authoritative plant taxonomy backbone (1.4M names, WCVP). Returns full classification, accepted/synonym status, native distribution, lifeform, synonyms, and child taxa. Search by name, family, or genus.
Scrapes the complete Kith product catalog from Shopify — prices, per-size availability, sale/markdown flags, and variant data for resale tracking and arbitrage.