Skip to content

Webzio-Extended

Bot in AI search and SEO
Moving Some facts here change over months and are rechecked monthly.

The training scraper Webzio-Extended is operated by Webz.io. It collects text in bulk for model training, which is a separate question from whether an AI product can cite the site today. It appears in server logs as webzio-extended.

What it feeds

Webzio-Extended is the second half of the Webz.io pair. Rather than fetching pages itself, it takes what the collector gathered and marks each item as usable or not usable for AI and machine learning training.[1] Webz.io says only the material tagged as permitted is sold on for AI purposes.[2]

Blocking and its consequences

The token is the AI-training opt out for the Webz.io corpus, separate from omgilibot, which is the collection side.[3] Webz.io states that its crawlers follow robots.txt exclusions.[4] Blocking it forfeits no answer-engine citations, because the output is a dataset rather than a live answer.

Verification

There is no documentation page or IP range list from Webz.io that we can point at for this token.

A user agent is a claim, not a proof: any client can send any name. Where an operator publishes IP ranges, a reverse lookup is the only way to tell a real visit from an impersonation.

Seen by Baseline

Requests, 30 days
...
Last 7 days
...
First seen
...
Last seen
...

Requests from this bot across sites tracked by Baseline, last 30 days. Aggregated over - sites, never reported per site.

References (4)
  1. Webz.io, Bot information
    Documentation Retrieved 2026-09-05
  2. Webz.io, Bot information
    Documentation Retrieved 2026-09-05
  3. Webz.io, Bot information
    Documentation Retrieved 2026-09-05
  4. Webz.io, Bot information
    Documentation Retrieved 2026-09-05

Last updated 2026-09-05. Written and maintained by Baseline Labs.

George
Online
0%