What is omgili? AI crawler guide
omgili is a crawler or fetcher associated with Webz.io. See source-graded user-agent, robots.txt, verification, and observation guidance.
What is omgili?
omgili is a documented Webz.io crawler or fetcher. Webz.io crawler for collecting web data sold through APIs and datasets.
Evidence status
| Field | Value |
|---|---|
| Evidence | Officially documented |
| Lifecycle | Active |
| Purpose | other |
| robots.txt posture | honors |
| Source checked | 2026-06-11 |
Documented user-agent
omgili
Allowing omgili only creates the possibility of retrieval. To find out whether your pages are selected, measure which answer engines actually cite your pages using a stable query set.
robots.txt allow example
User-agent: omgili Allow: /
robots.txt block example
User-agent: omgili Disallow: /
What the rule can and cannot do
Robots.txt expresses an access policy to compliant automated crawlers. It does not authenticate the sender, remove content already collected, or guarantee that a model will use or not use content.
How to verify a request
- Match the complete documented user-agent where one exists. A match is only a clue.
- Look for operator-published IP ranges, reverse DNS, or signed-request guidance. None is attached to this record.
- Keep the observed request, operator documentation, and any inference as separate fields.
JavaScript behavior
The cited operator material does not verify JavaScript rendering. Serve useful HTML before client-side JavaScript where possible.
Related bots
- Brightbot: Also tracked as a general crawler.
- Amazonbot: Also tracked as a general crawler.
- ImagesiftBot: Also tracked as a general crawler.
- omgilibot: Another Webz.io general crawler to compare.
- Diffbot: Also tracked as a general crawler.
- GoogleOther: Also tracked as a general crawler.
- GoogleOther-Image: Also tracked as a general crawler.
- GoogleOther-Video: Also tracked as a general crawler.
- Panscient: Also tracked as a general crawler.
- Robots.txt: Robots.txt is the control file used to allow or block omgili.
- AI Crawlers: omgili is a concrete crawler example for this concept.
Frequently Asked Questions
What is omgili?
omgili is a officially documented crawler or fetcher record associated with Webz.io.
What user-agent does omgili use?
omgili. Check the evidence status before attributing a matching request.
Can omgili be blocked in robots.txt?
The record identifies omgili as the token to review. Robots.txt is a request policy, not proof of model use or non-use.
How can I verify a omgili request?
Treat the user-agent as a clue and look for operator-published IP, reverse-DNS, or signature guidance.
Data & Sources
- Webz.io documentation - Primary source for omgili crawler details.