Scrape NYC building violations, permits, and property data from 7 open city datasets (DOB violations, ECB violations, HPD housing violations, permits, job filings, PLUTO property profile, LL84 energy benchmarking) joined on BIN/BBL into a single building-risk profile per address.
Scrapes OFA breed-level disease-prevalence statistics for hip dysplasia, elbow dysplasia, cardiac disease, patellar luxation, thyroid, eyes, and DNA tests. Returns per-breed evaluation counts and percentage breakdowns — used by pet-insurance actuaries and AKC breeder compliance auditors.
Scrape ebook and audiobook availability from any public library's OverDrive/Libby catalog. Provide library slugs and search queries — get real-time copy counts, holds, wait times, and full metadata. The digital-circulation signal publishers and authors can't source elsewhere.
Scrape board-certified plastic surgeons from the ASPS find-a-surgeon directory. Returns name, board certs, practice address, phone, email, website, languages, and lat/lon — structured for B2B lead-gen in aesthetic medicine.
Scrape the QueryTracker literary agent directory — the #1 querying-author platform. Extracts agent name, agency, genres, query status, query method, social links, and last-updated date from every public agent profile. Ideal for query-CRM tools and literary data pipelines.
Scrape ranked startup lists from SeedTable — startups across 4,000+ city lists worldwide. Returns startup name, industry, location, profile URL, and logo for every ranked company, and filters to just the cities you care about.
Company org charts, corporate email-format patterns, competitors, subsidiaries and workforce tenure/turnover data pulled from SignalHire company profiles. No login or account needed. Covers the full published directory, A-Z and 1-9.
Scrape public profile metadata from Snapchat user profiles. Provide usernames or profile URLs and get display name, subscriber count, bio, avatar URL, and external links — one record per username.
Generate realistic e-commerce test data with interconnected products, customers, orders, and reviews. Features referential integrity, realistic distributions, temporal coherence, industry presets, and deterministic seed mode.
Bulk TikTok creator data by username or URL: followers, bio, verification, total likes and region, plus each creator's videos with views, likes, comments, shares, saves, hashtags, sound and play/download URLs. Comments optional.
Scrape business profiles and reviews from Trustpilot. Extract trust scores, ratings, star distributions, review text, and company details. Supports search queries, direct URLs, and category browsing.
Scrape software reviews and product data from TrustRadius. Extract ratings, trScores, reviewer details, pros/cons, and company metadata for competitive intelligence and market research.
Scrape full restaurant menus from UberEats. Extract restaurant info, all menu sections, items with prices, descriptions, and images.
Scrapes Ukraine ProZorro public-procurement contracts in reconstruction CPV / DK021 categories (construction, engineering, demolition, capital repair). Emits buyer, supplier, EDRPOU codes, contract value, dates, item descriptions. No auth.
Scrape federal job listings from USAJOBS.gov. Search by keyword, location, agency, salary, GS grade, work schedule, remote/telework, and hiring path. Extracts title, agency, department, location, salary, grade, open/close dates, and job URL.
Daily delta of newly registered domains from the whoisds.com free NRD feed, enriched with DNS (A/MX/NS), Certificate Transparency presence, IDN punycode decoding, and keyword/typosquat matching against a brand watch-list. Fully passive.
Scrapes the Lady Bird Johnson Wildflower Center NPIN — the canonical US native-plant database with ~9,000 species. Returns scientific name, common names, USDA symbol, growing conditions, bloom data, wildlife value, pollinator and butterfly larval-host data, and commercial availability.
Scrape Philippines transit: inter-island ferries (2GO, Montenegro, Oceanjet, FastCat, Starlite) with cabin-class breakdown and RORO vehicle fares in PHP, PNR/MRT-3/LRT-1/LRT-2 rail, PITX provincial bus, and Cebu Pacific domestic flights.
Scrapes the ACE public exercise library. Extracts professionally-authored exercises with step-by-step instructions, body parts, equipment, difficulty, images, and attribution. For fitness apps, RAG systems, and health pipelines that require provenance-verified exercise content.
Scrapes ai-bot.cn — the definitive curated China AI tool directory — into a clean dataset of 1,000+ tools with names, categories, descriptions, logos, and external links. Covers domestic (China) and overseas AI tools with origin classification.
Short-term and vacation rental listing data, by location. Returns nightly price, property and room type, amenities, host and superhost status, rating, review count, guest capacity, photos and coordinates. Optional dates and guest count. For rental market research, pricing benchmarks and host lists.
Track Amazon Best Sellers rankings across any category. Collects product name, rank, ASIN, price, rating, review count, and image URL from any bestsellers category page.
Scrape US intercity bus schedules and fares: Megabus and FlixBus (Greyhound/FlixBus network) with per-carrier trip IDs, origin/destination cities, departure/arrival times, fares, seat availability, amenities, and booking links. Covers 734+ city pairs and 500+ routes across the continental US.
Extract German job listings from Arbeitsagentur.de (Bundesagentur für Arbeit), Germany's federal employment agency. Search by keyword, location, or browse all open positions. Returns job title, employer, location, contract type, posted date, and job URL.