Scrape Reddit posts, comments, search results, and user profiles. No API keys or browser needed. Supports 4 modes: subreddit posts (hot/new/top/rising), Reddit search, user profiles, and full comment trees. Fast, lightweight HTTP-based scraping with built-in rate limiting and retry logic.
Scrape LinkedIn job listings without API keys, login, or browser. Extract titles, companies, locations, salaries, descriptions, and more from public job search pages. Fast HTTP-based scraping with rich filters.
Find healthcare provider emails and contacts from NPI registry. Generate sales leads with doctor emails, LinkedIn profiles, practice websites. No API key.
Scrape Google Maps places, reviews, and business leads. No API keys, no login. Headless Chrome with parallel batch keywords × locations, optional phone/website/hours enrichment, and email + social extraction. Pairs with linkedin-jobs-scraper for full lead pipelines.
Search 35M+ medical citations from PubMed/MEDLINE. Extract articles, abstracts, authors, MeSH terms, and citations for research, competitive intelligence, or AI/RAG pipelines. No API key required.
Apify Actor that writes embedding vectors to Pinecone or Qdrant vector databases. Chains directly with RAG Embedding Generator output or accepts raw vectors with metadata. Handles batching, retries, collection creation, metadata mapping, and ID generation. Bring your own vector DB API key.
Search MILLIONS of academic papers from Semantic Scholar and arXiv by keyword, DOI, or citation graph. Returns titles, authors, abstracts, citation counts, and open access PDFs as clean JSON. Works as an MCP tool for AI agents.
Search concerts, sports, theater, and live events from Ticketmaster by keyword, location, date, or genre. Returns ticket prices, venue details, presale dates, seat maps, and direct purchase links as clean JSON. Works as an MCP tool for AI agents.
Turn raw text, Markdown, or Apify datasets into token-perfect RAG chunks with deterministic IDs, source metadata, and a billing-ready summary—ready for embeddings or vector DBs without extra glue code.
Generate vector embeddings from text or chunked datasets using OpenAI or Cohere. Chains with RAG Content Chunker for end-to-end RAG pipelines. Outputs raw vectors ready for any vector database.
Find emails, phone numbers and social profiles on any list of websites. Reads contact, about, team and imprint pages in many languages, decodes Cloudflare-protected and [at]/[dot] emails, validates phones, reads schema.org data, optional MX email check. One row per website.
One-click RAG pipeline: chunks text, generates embeddings, and stores vectors in Pinecone or Qdrant. Provide your content and API keys -- the orchestrator handles the rest.
Search U.S. federal contract opportunities, awards, and agencies from SAM.gov. Filter by keyword, NAICS code, set-aside type, state, agency, and more. Returns structured data including contacts, deadlines, award amounts, and direct SAM.gov links. Requires a free SAM.gov API key.
Extract FDA Orange Book data — drug patent expirations, exclusivity periods, generic equivalents, and therapeutic equivalence ratings. No API key required.
Extract structured pricing data from any SaaS company's public pricing page. The dedicated SaaS pricing intelligence tool on the Apify Store.
Scrape the Apple App Store: search apps by keyword, get full app details by ID or URL, pull up to 500 reviews per app per country, or grab the Top Free, Top Paid and Top Grossing charts by category. Any country. No API key.
Extract FDA drug label data -- indications, dosages, warnings, black box alerts, drug interactions, and more. Search by drug name, active ingredient, manufacturer, or browse the full openFDA database. No API key required.
Turn seed keywords into hundreds of real search suggestions from Google, YouTube, Google Shopping, Amazon and Bing. Alphabet, question and preposition expansion, recursive depth, relevance scores, any language and country. For SEO, PPC, content and product research.
Scrape Google News by keyword or topic in any language and country. Real publisher URLs (not news.google.com redirects), source, date and related coverage. Go past Google's 100-result cap with a date range: one feed per day.
Scrape Google Play: search apps, full app details (installs, ratings histogram, developer contact, version), unlimited reviews with star and date filters, and Top Free, Paid and Grossing charts by category. Any country and language. No API key.
Find Reddit posts with buying intent: people asking for a tool like yours, unhappy with a competitor, or ready to switch. Delivered scored, ranked, and with a suggested reply on each. No API key, no login.
Who owns what in orbit. Every satellite operator with its active fleet, parent company after mergers, country, HQ, orbits and makers. Find new entrants for lead lists, or any satellite's ownership chain. Built on GCAT.
Scrape Substack newsletters: every post with title, date, engagement, paywall status and full text for free posts, plus comments, publication details and subscriber counts. Or search all of Substack by keyword. No login.
See what any website is built with: CMS, ecommerce, analytics, ad pixels, JS frameworks, CDN, hosting, payments, live chat and email provider. 6,000+ technology fingerprints plus DNS, with versions and evidence. BuiltWith-style lookups in bulk.