Crawl toxic chemical release data from the EPA TRI via the Envirofacts API. Extract facility details, chemical names, release quantities by media (air, water, land), coordinates, and carcinogen flags. Filter by state, chemical, year, and facility.
Scrape the EU Transparency Register: organisations, financial disclosure, clients, accredited lobbyists, and Commissioner meetings. Ideal for public-affairs intelligence, ESG screening, political-risk research, and compliance due diligence.
Scrape company listings from Europages, Europe's largest B2B supplier directory. Search by product, service, country, and company type. Extract company name, address, phone, website, products, certifications, and employee range.
Scrape expert witness profiles from Experts.com — the largest open expert witness directory. Extracts contact info, specialty, credentials, categories, and biography for thousands of expert witnesses across all legal specialties.
Scrapes the FAA Office of Commercial Space Transportation (AST) website for commercial space launch license actions, program pages, and authorization announcements. Returns structured records from FAA press releases and stakeholder engagement pages covering licensed launch and reentry operations.
Scrape Failory's live startups directory — 14,000+ startups across 267+ country, city, and industry facet pages. Extracts startup name, website URL, industry, year founded, funding amount, funding round, and facet label. Ideal for lead generation, VC research, and competitive intelligence.
Scrapes FCC IBFS satellite service filings from fcc.report. Covers earth-station (SES) and space-station (SAT) applications with applicant, status, frequency bands, attached document URLs, and public notices. Supports date-range filtering for incremental runs.
Crawl campaign finance data from the FEC. Extract candidates with financial summaries, committees/PACs, and individual contributions. Filter by year, state, party, and office. For political researchers, journalists, and compliance teams.
Scrapes FishBase — 35,000+ fish species with taxonomy, ecology, size/weight, trophic level, IUCN status, game-fish flag, depth range, and common names.
Scrapes Flickr's camera-finder popularity dataset — photo upload counts, photographer user counts, and popularity scores for 600+ camera models across 50+ brands. Uses Flickr's internal camerafinder API to deliver the canonical real-world camera adoption signal.
Scrapes ~800 franchise brands from FranchiseDirect.com. Extracts franchise fee, minimum cash required, total investment range, territory availability, and brand description across all industry categories.
Extract book data from Goodreads: titles, authors, ratings, reviews, genres, ISBN, pages, format, publication date, awards, and more. Accepts book or author URLs as input.
Search and scrape US court opinions from Google Scholar's case law database. Filter by keywords, courts, and date ranges to collect case names, citations, courts, dates, snippets, "Cited by" counts, and links. Optionally fetch full opinion text from detail pages.
Grants.gov scraper for federal grant opportunities (posted, forecasted, closed, archived). Crawl titles, agencies, CFDA numbers, funding amounts, eligibility, close dates, and grantor contacts. Filter by agency, category, instrument, eligibility, CFDA, keyword, or posted-date window.
Scrapes race info, packages, registration details, and results history from the official Great Wall Marathon site (great-wall-marathon.com). Returns structured records for the current race edition, package pricing, and historical result editions.
Extracts the worldwide directory of Herman Miller authorized dealers, retailers, showrooms, and resellers from the public dealer locator API. Returns name, address, contacts, certification level, product lines, and business categories for every location globally.
Scrapes HobbyLink Japan (hlj.com), the largest English-language Japanese hobby shop, for Gunpla and anime figure catalog data: names, prices in JPY, release status, manufacturer, series, scale, JAN barcodes, and image URLs. Category backfill, new-release weekly, and pre-order modes.
Extract home-services pro business listings from HomeAdvisor's directory by trade and location: business name, rating, review count, years in business, screening status, and service area.
Compare hotel prices across Booking.com, Expedia, Hotels.com, and Priceline from a single search. Collect nightly rates, guest ratings, star ratings, amenities, and photos for any destination. Perfect for travel market research and rate intelligence.
Scrape AI/ML model metadata from the HuggingFace Hub. Extract model names, task types, download counts, likes, libraries, authors, tags, licenses, model sizes, and model card excerpts. Filter by task type, library, author, and search query.
Aggregates 6 major HVAC OEM dealer networks (Trane, Lennox, Carrier, Rheem, Goodman, American Standard) into one structured feed with manufacturer certification tier attached.
Scrape property listings (apartments, houses, for-sale and rental) from ImmobilienScout24.de. Returns price, living space, rooms, address, realtor, features, and images per listing.
Scrapes Brazil's INPE satellite imagery catalog (CBERS-4, CBERS-4A, Amazonia-1) via the STAC API. Supports filtering by collection, bbox, date range, and cloud cover. Returns scene metadata including satellite, sensor, path/row, acquisition date, cloud cover, sun angles, thumbnail and download URLs.
Search the iTunes/Apple Music catalog by keyword, artist, or album; look up tracks by Apple ID; and fetch Top Charts rankings for any country. Metadata and 30-second preview URLs only — no audio downloads.