Scrapes the ACE public exercise library. Extracts professionally-authored exercises with step-by-step instructions, body parts, equipment, difficulty, images, and attribution. For fitness apps, RAG systems, and health pipelines that require provenance-verified exercise content.
Scrape the Adobe Commerce (Magento) Marketplace. Extract extension names, vendors, USD list prices, edition and Magento-version compatibility, ratings, reviews, and install metrics. Built for Magento agencies, extension ISVs doing competitive pricing, and ecommerce platform researchers.
Extracts child well-being indicator data from the Annie E. Casey Foundation KIDS COUNT Data Center. Covers hundreds of indicators across all 50 US states and national level, multiple years. Fields: indicator name, location, year, data format, and numeric value.
Scrapes ai-bot.cn — the definitive curated China AI tool directory — into a clean dataset of 1,000+ tools with names, categories, descriptions, logos, and external links. Covers domestic (China) and overseas AI tools with origin classification.
Short-term and vacation rental listing data, by location. Returns nightly price, property and room type, amenities, host and superhost status, rating, review count, guest capacity, photos and coordinates. Optional dates and guest count. For rental market research, pricing benchmarks and host lists.
Track Amazon Best Sellers rankings across any category. Collects product name, rank, ASIN, price, rating, review count, and image URL from any bestsellers category page.
Scrape US intercity bus schedules and fares: Megabus and FlixBus (Greyhound/FlixBus network) with per-carrier trip IDs, origin/destination cities, departure/arrival times, fares, seat availability, amenities, and booking links. Covers 734+ city pairs and 500+ routes across the continental US.
Scrapes Anytime Fitness club locations from their sitemap — covers 3,000+ US and international clubs. Extracts club name, address, contact details, coordinates, hours, and amenities from each location page.
Extract German job listings from Arbeitsagentur.de (Bundesagentur für Arbeit), Germany's federal employment agency. Search by keyword, location, or browse all open positions. Returns job title, employer, location, contract type, posted date, and job URL.
Scrapes California State Bar disciplinary actions — disbarments, suspensions, probations, censures. One row per action with attorney name, bar number, sanction type, effective date, violations, and profile link.
Scrape property auction listings from Auction.com. Extract address, price, auction date, status, beds/baths, square footage, and more. Filter by US state or price range. Ideal for real estate investors, lead generation, and market research.
Scrape Audubon's "Plants for Birds" database by US ZIP code — returns native plants ranked for bird habitat value with bird species attracted, plant type, and wildlife resources provided. Unique zip → native-plant → bird-ecology join not available elsewhere as a structured data feed.
Scrape the Auth0 Marketplace. Extract integration names, vendors, categories, feature types, tags, and descriptions across social connections, MFA, SSO, and security integrations. Built for identity/security ISVs and CIAM competitive intelligence.
Scrape Babylist's product catalog: product details, pricing, ratings, multi-retailer links, editorial badges, and category data from the leading US baby registry platform. Supports sitemap-driven full crawl, category browsing, editorial best-of list extraction, and direct product URL lookup.
Scrape Behance project metadata including title, owner, tags, engagement stats, and cover image URL. Search by keyword — illustrations, paintings, branding, and more.
Scrapes B&H Photo's Pro Audio catalog: microphones, studio monitors, audio interfaces, mixers, headphones, signal processors, and podcast gear. Extracts SKU, price, stock status, ratings, specs, and features. Supports used/B-stock inventory.
Scrape Japan's official court auction database (BIT / 不動産競売物件情報サイト) — all 50 district courts. Extracts case numbers, addresses, property types, areas, bid prices, deadlines, and 3-point-set PDF links for below-market foreclosure listings.
Scrape book series reading order data from bookseriesinorder.com. Returns publication order, chronological order, series positions, and book metadata for thousands of authors and series.
Scrape property listings from Caixa Econômica Federal's real estate auction portal. Filter by state, city, modality and minimum discount. Returns appraisal value, discount %, address and edital link.
Scrape individual C2C marketplace listings from Carousell — Southeast Asia's largest C2C platform. Extract listing details, pricing, seller profile (username, rating, review count), condition, images, and engagement metrics. Supports all 8 Carousell country domains.
Scrapes sanctioned cycling events and state federations from the Confederação Brasileira de Ciclismo (CBC). Returns event details, dates, locations, disciplines, and federation contact information.
Scrape new and pre-owned luxury watch listings from Chronext — Europe's leading multi-currency watch retailer. Captures brand, model, reference number, price (EUR/GBP/CHF/USD), condition, provenance, and full specifications for cross-border arbitrage and price tracking.
Scrape church listings from ChurchFinder.com — US churches across all 50 states with name, address, city, state, ZIP, phone, denomination, service times, and description. Search by state or provide direct URLs. Ideal for religious org research and directory building.
Scrapes normalized GPU cloud pricing from Runpod and Vast.ai. Returns per-GPU-hour prices with hardware specs and availability across community and secure tiers. Useful for AI training cost comparison, arbitrage monitoring, and GPU price trend analysis.