ScrubberDuck navigating crawler feeds and research data

Scrubnet

Verified crawler observations and live request data

Explore Live Logs

What is Scrubnet?

Scrubnet is an independent public observatory for understanding how search crawlers and AI bots behave. We compile authorised website content into machine-readable feeds, record how verified bots discover and revisit them, and publish the resulting data, observations and practical tools.

Why study crawlers?

Who It’s For

Our Principles

What Scrubnet brings together

Scrubnet combines a growing set of machine-readable feeds, a public crawler log dashboard, observational reports and free tools. Together they help us move beyond crawler speculation and build a clearer picture of how verified crawlers interact with public content.

Meet ScrubberDuck

ScrubberDuck is our lightweight, robots-aware feed compiler. It collects authorised public content and creates the Scrubnet feeds used to observe discovery, formats, freshness signals and recrawl behaviour.

It’s designed to minimise load, avoid unnecessary requests, and respect robots.txt.

Illustration of ScrubberDuck, the lightweight crawler used by Scrubnet

If you see ScrubberDuck in your logs, it means your site is contributing authorised content to the Scrubnet observatory.

User-Agent: ScrubberDuck/1.0 (+https://scrubnet.org)

How the research works

We add authorised participating websites, fetch their public pages efficiently and publish consistent machine-readable feeds. We then monitor which verified and unverified bots request those feeds, when they return, which formats they choose and how requests relate to content changes.

The observations feed into descriptive reports, technical SEO and GEO guidance, and tools such as SEO Scrubbox. Adding more varied sites broadens the observable content base and makes the dataset more representative of the participating sites.

Participation is free: add a website with up to 50,000 public URLs. There are no visibility or ranking guarantees.

Crawlers we monitor

The live dashboard records recognised search, AI and archive crawlers requesting Scrubnet feeds, including:

Get Involved

Add a site to expand the observational dataset, explore the live logs, or collaborate with us on a crawler analysis or case study.

Or reach out at contact@scrubnet.org

Inspect crawler-facing SEO signals with our free Chrome extension

SEO Scrubbox

A Chrome extension for technical SEO and crawler diagnostics. Compare view-source vs rendered signals, spot canonical drift, validate JSON-LD, audit sitemaps, hreflang, redirects, and crawl signals without leaving the page.

Member of Manchester Digital Badge