omgilibot
The Webz.io crawler, harvesting forums, reviews and news into datasets sold for research and AI training. Appears in logs under the older omgili name.
What it feeds
Blocking and its consequences
Webz.io states that its crawlers honour robots.txt exclusions.[3] Because the output is a dataset sold to third parties rather than a live answer engine, blocking costs no citations: it removes the site from a training and research corpus.[4]
History
Omgili began as a forum search engine and the crawler kept its name after Webz.io took over the data business, which is why a modern log line can carry a token whose brand no longer exists.[1] The current pairing splits the work between a collector and Webzio-Extended, which does the permission tagging.[2]
Verification
Operator documentation: https://webz.io/bot.html
A user agent is a claim, not a proof: any client can send any name. Where an operator publishes IP ranges, a reverse lookup is the only way to tell a real visit from an impersonation.
Seen by Baseline
Requests from this bot across sites tracked by Baseline, last 30 days. Aggregated over - sites, never reported per site.
References (4)
Last updated 2026-09-05. Written and maintained by Baseline Labs.