Find broken links (404s, dead domains, server errors) on any list of pages. Checks every link on each page, follows redirects, retries temporary errors, and can include images, scripts and stylesheets. One row per link with status code and link text.
Check any list of domains for email authentication and spoofing protection: SPF (including +all and the 10-lookup limit), DMARC policy and reporting, DKIM keys on 30 common selectors, MX and mail provider, MTA-STS, TLS-RPT and BIMI. A-F grade and a plain list of issues per domain.
Audit web pages for on-page SEO in bulk: title and meta description length, headings, canonical, noindex, mobile viewport, image alt text, links, word count, Open Graph, schema.org, HTTPS and speed. Each page gets a 0-100 score and a plain list of issues to fix.
Get every page URL a website lists in its XML sitemaps, with last-modified date, change frequency, priority and image count. Finds sitemaps via robots.txt, follows sitemap indexes, reads .xml.gz and text sitemaps.
Give it VINs, get specs, safety recalls, owner complaints and crash-test ratings from official US government data. Built for AI agents and repair, insurance and used-car workflows.
Turn any web page into clean Markdown for LLMs, RAG and AI agents: main content only (no menus, ads or footers), absolute links and images, tables kept, title/author/date metadata, word and token counts, optional chunks. Respects robots.txt. Pay per page, errors free.
Find the CMS, e-commerce platform, frameworks, analytics, ad pixels, hosting, CDN, payments and chat tools behind any website. Bulk domains, version numbers, evidence for each match.
Get every open job from company career pages that run on Greenhouse, Lever, Ashby, Workable, SmartRecruiters or Recruitee, in one clean format: title, location, remote, department, type, salary where published, posted date, apply link and full description. Paste careers URLs or just company names.
Clean an email list in bulk: syntax check, does the domain exist and accept mail (MX records), disposable/throwaway providers, role addresses like info@, free providers like Gmail, and typo suggestions (gmial.com). No emails are sent and mailboxes are not probed.
Pull the machine-readable data out of any web page: schema.org JSON-LD, microdata types, Open Graph and Twitter cards, product name/price/stock/rating, organisation contact details, articles, breadcrumbs and FAQs. Bulk URLs, errors free.