Scrape the latest AI, ML, and data science job listings from foorilla.com/hiring (formerly aijobs.net). Extracts salary range, seniority, years of experience, remote policy, skills, education, tasks and apply URL. Covers the freshest ~105 public postings.
Amtrak's full US passenger rail network: every route (Acela, Northeast Regional, Silver Meteor, Coast Starlight and ~50 others), 1,000+ stations with codes, state, timezone and amenity flags (wheelchair, staffed, QuikTrak), and current service alerts.
Scrape Australian Business Register (ABR) entities by name search or direct ABN lookup. Returns ABN, ACN, entity name, type, ABN/GST status, main business location, business names, trading names, and full status history — the complete KYB picture from Australia's free authoritative register.
Scrapes retreat center profiles and reviews from AyaAdvisors, the leading directory for ayahuasca retreats. Covers legal-jurisdiction centers across Peru, Costa Rica, Brazil, Colombia, and Ecuador with ratings, reviews, pricing, and contact data.
BBB Business Scraper for Better Business Bureau listings: search bbb.org by category and location across the US and Canada and extract business name, BBB rating, accreditation status, phone, address, website, emails and contacts. Lead generation from the BBB directory.
Scrape attorney profiles from BestLawyers.com, the oldest peer-reviewed legal ranking. Extract names, firms, practice areas, contact info, education, awards, and recognition years. Filter by US state and city. Covers 150+ practice areas across 75+ countries.
Scrape auto parts from CarParts.com. Search by keyword, part number, brand, or category. Extracts part number, brand, price, shipping, core price, warranty, vehicle compatibility, OE cross-references, and detailed specs. Ideal for automotive market research and price monitoring.
Scrape recipes from Chefkoch.de — Germany's largest recipe community with 380,000+ recipes. Extracts full structured data including ingredients, instructions, nutrition, ratings, and Chefkoch-specific fields (Schwierigkeitsgrad, Portionen). Supports bulk export and filtering by keyword or category.
Crawl US court opinions from CourtListener (Free Law Project). 8M+ decisions from 3,350+ federal, state, and appellate courts. Filter by court, judge, date range, citation, or keyword. Returns case metadata, citations, precedential status, and opinion excerpts.
Licensed care facilities from four state regulators. California covers every Community Care Licensing licence — daycare, preschool, assisted living, adult residential, foster and group homes, home care, adoption. Florida, Texas and New York cover child care. Search by ZIP, city, county or name.
Scrape global job postings from Hiring Cafe (hiringcafe.com) — an AI-powered job search engine.
Extract job title, company, location, salary, remote policy, seniority level, visa sponsorship,
commitment type, and more from thousands of live listings.
Search Sweden's hitta.se public person directory by name and get age, phone, and street address, plus municipality, county, and GPS coordinates, for private individuals. Built for debt-collection, tenant screening, insurance investigation, and B2C lead-list building.
Scrape the full IMSLP public-domain score catalog — 230k+ works across 24k composers, with file URLs, copyright tags, and work metadata via the IMSLP worklist API and MediaWiki API.
Crawl Indian court judgments and orders from the Supreme Court, all 24 High Courts, Tribunals, and central laws. Filter by court, date range, judge, citation, or keyword. Optional full judgment text. Requires user-supplied Indian Kanoon API token (paid - apply at api.indiankanoon.org).
Scrape Japan rail timetables and fares from Yahoo! Transit. Covers JR Group, every Shinkansen line, Tokyo Metro, Toei, and major private rail (Odakyu, Tokyu, Keio, Hankyu, Kintetsu, etc). Returns IC fare, transfers, segments, train type, JR Pass eligibility. Hyperdia replacement.
Fetch Disney Lorcana card data from the Lorcast open API. Supports all_cards (full catalog), set (single set), search (free-text query), and card_ids (specific Lorcast IDs) modes. No API key required. Returns ink, cost, keywords, prices, images, and legalities.
Scrape property listings from Mexico's top portals — Inmuebles24 and Vivanuncios. Filter by operation (sale, rent), property type, and location. Returns price, area, bedrooms, and photos. Built for US investors, PropTech pipelines, and market analysts.
Scrapes the Moonshot Kimi API pricing catalog for all current models — Kimi K3 (flagship), Kimi K2.7 Code, and Kimi K2.6. Returns per-model input/output token prices (CNY), cache-hit pricing, context window size, modality, and a USD equivalent using the day-of-run exchange rate.
Scrape MTGTop8 -- 20 years of Magic: The Gathering tournament data. Extract events, decklists, archetypes, and meta-share metrics. Five modes: events by format, single event, single deck, top-cards meta rollup, or archetype rollup.
Crawl 250K+ CVE records from the NIST National Vulnerability Database. Extract CVSS v3.1 scores, severity ratings, attack vectors, affected products (CPE), CWE weaknesses, and exploit/patch status. Filter by keyword, severity, vendor, and date range.
Scrape distressed real estate listings from RealtyTrac — pre-foreclosures, active foreclosures, bank-owned (REO) properties, and sheriff sales. Filter by US state or city. Returns address, beds/baths/sqft, estimated value, listing status, and detail page URL.
Scrapes 50 000+ live listings from Rebag, a top-3 US luxury consignor. Outputs brand, price, condition, hardware, and native price-band facets for Hermès, Chanel, LV, Cartier and 100+ more brands. Ready to join with other consignor datasets for cross-platform price analysis.
Crawl SEC EDGAR company data and filings. Extract tickers, CIK, SIC codes, addresses, and filing records. Search by ticker, company name, CIK, or SIC code. 800K+ companies, 12M+ filings.
Scrapes live Steam charts — most played, top sellers, new releases, specials — with per-game live player counts and an optional premium add-on for each game's official release date. Returns time-stamped rows for trend pipelines, launch benchmarking, and market intelligence.