Scrape GSA eLibrary — all Schedule contract holders across every Special Item Number (SIN). Captures name, contract number, SIN, contact details, address, SAM UEI, socioeconomic indicators (SDVOSB/WOSB/HUBZone), and contract dates.
Digital product creator profiles from Gumroad and Payhip, two of the largest independent seller platforms. Returns creator name, bio, avatar, social links, and a full product-and-price list per storefront — across the whole creator base on each platform, not one search term.
Enumerates the active business-license roll of HdL Companies municipal portals (Pomona, Hayward, El Cajon, Tustin and other reachable HdL cities) by business-type code, returning account number, business name, license status, expiry, and address for every license on file.
Local service pro directory scraped from HomeGuide across every US state, city, and trade — business name, phone, website, rating, reviews, years in business, and service area for plumbers, electricians, HVAC, roofers, landscapers, cleaners, pest control, movers, and more.
Extract live salvage-vehicle auction listings from IAAI's public search. Search by keyword or filter by state. Returns stock number, partial VIN, year/make/model, damage type, odometer, title type, run-and-drive status, sale date, branch location, and more.
Search India's largest free case-law and legislation index by keyword. Get structured judgment metadata: court, citation, decision date, bench, full text, and citation counts for legal research and litigation analysis.
Search Australian trademark applications and registrations from IP Australia's Trade Marks Online Search. Filter by name, owner, application number, Nice class, and status to pull mark text, owner, goods & services, filing/registration dates, and current status.
Download and parse the IRS Tax Exempt Organization auto-revocation list. Returns all nonprofits that lost tax-exempt status for non-filing, including reinstatement records.
Scrapes the public JailbreakBench leaderboard tracking attack-success-rate for jailbreak techniques (PAIR, GCG, AIM, and more) against open- and closed-source LLMs, with and without defenses (SmoothLLM, perplexity filter, etc). Snapshot each run to track technique-vs-model ASR movement over time.
Job postings across Malaysia, Singapore, the Philippines, Indonesia, Hong Kong and Thailand — sourced from JobStreet and JobsDB (SEEK Asia). Get titles, companies, salaries, locations, classifications and full descriptions in one dataset.
Extract Kavak's owned used-car inventory across Mexico, Brazil, Argentina and Chile: price, financing, mileage, spec, hub location and availability for every listed vehicle, normalised into one cross-country schema.
Scrape Kickstarter campaigns across the whole site — all 15 categories and 154 subcategories — with funding goal, amount pledged, percent funded, backer count, creator history, launch date and deadline. Pick any mix of categories, or crawl everything.
Scrape classified listings from Kleinanzeigen.de (formerly eBay Kleinanzeigen). Extract listing details including title, price, description, location, seller info, images, and category-specific attributes.
Extract overseas property listings and the estate agencies behind them from Kyero across Spain, Portugal, France, Italy, Andorra, Greece, Cyprus and Turkey.
Normalizes the scattered public GitHub collections of leaked/published AI system prompts (ChatGPT, Claude, Cursor, Devin, v0, Perplexity, and more) into one deduplicated dataset with product, vendor, version, and leak-date fields. Passive: reads only already-public repositories.
Scrape classified ads from leboncoin.fr — title, price, location, images, attributes, and owner type. Pass any category or search URL.
Scrape auction property listings from leilaoimovel.com.br — Brazil's largest real estate auction portal. Extracts sale price, appraisal value, discount percentage, modality, closing date, bank, property type, address, and more. Filter by state, city, property type, and bank.
Scrapes employer reviews, ratings, and salary data from en-hyouban.com (エン カイシャの評判) — Japan's #2 employer-review platform. Extracts JSON-LD EmployerAggregateRating, per-dimension breakdowns, crowd-sourced salary, and employee reviews by URL or sitemap.
Scrape LinkedIn Ad Library ads by keyword, company ID, or country: advertiser name, ad creative, and the full EU DSA disclosure block — total and per-country impressions, audience targeting, and the legal entity paying for each ad.
Scrape public LinkedIn Services Marketplace listings across 200+ service categories (accountants, developers, designers, lawyers, and more). Returns provider name, category, headline, description, skills, location, rating, and review count. Auth-less, no login needed.
Swiss business directory data: company names, addresses, phone numbers, emails, ratings, and opening hours for the Swiss SMB market. Filter by industry vertical. Structured JSON output ready for lead lists or CRM import.
Validate any NANP phone number and get back the real carrier of record, ILEC, rate centre, and a wireless-vs-landline flag from public numbering-plan data — plus local time and a TCPA calling-window check, not a format parse.
Scrape used heavy equipment listings from MachineryTrader.com across every major equipment category (excavators, skid steers, dozers, cranes, loaders, and more). Get pricing, full specs, dealer contact info, location, and photos for every for-sale listing.
Cross-ecosystem feed of confirmed malicious/typosquatted packages across npm, PyPI, crates.io, Go, Maven, NuGet, Packagist and RubyGems, sourced from the OSSF malicious-packages dataset (also feeds OSV.dev). Category taxonomy plus typosquat-target linkage for CI gating. Passive read only.