Skip to content

omgili

Bot in AI search and SEO
Moving Some facts here change over months and are rechecked monthly.

The training scraper omgili is operated by Webz.io. It collects text in bulk for model training, which is a separate question from whether an AI product can cite the site today. It is a robots.txt token only and never appears in server logs.

What it feeds

The token omgili is the robots.txt name for the Webz.io collection pipeline, kept from the crawler's earlier identity.[1] The pages it gathers end up in datasets sold for research and for AI training, with a permission tag attached.[2]

Blocking and its consequences

Webz.io states that its crawlers adhere to robots.txt exclusions, so the token is worth writing out in full alongside omgilibot.[3] The company also says it only supplies data marked as permitted for AI use.[4]

Verification

There is no documentation page or IP range list from Webz.io that we can point at for this token.

A user agent is a claim, not a proof: any client can send any name. Where an operator publishes IP ranges, a reverse lookup is the only way to tell a real visit from an impersonation.

Seen by Baseline

Requests, 30 days
...
Last 7 days
...
First seen
...
Last seen
...

Requests from this bot across sites tracked by Baseline, last 30 days. Aggregated over - sites, never reported per site.

References (4)
  1. Webz.io, Bot information
    Documentation Retrieved 2026-09-05
  2. Webz.io, Bot information
    Documentation Retrieved 2026-09-05
  3. Webz.io, Bot information
    Documentation Retrieved 2026-09-05
  4. Webz.io, Bot information
    Documentation Retrieved 2026-09-05

Last updated 2026-09-05. Written and maintained by Baseline Labs.

George
Online
0%