What is Funnelback?
Funnelback crawls the websites and data repositories configured for a Squiz enterprise search collection. It turns the retrieved documents into an index used by an organization's own website search or internal search. This is a scoped collection workflow, not a crawler building a general public web index.
Collection administrators decide which hosts and URL patterns belong in a crawl. Funnelback can follow links within that scope and, when the relevant option is enabled, read sitemap locations from robots.txt. A request in your logs usually means that a configured collection is discovering or refreshing material for its search results.
Squiz's crawler documentation says Funnelback honors the original robots.txt standard by default. Its support is narrower than that of some modern crawlers: Allow rules and path wildcards are not supported. An administrator can also configure a collection to ignore robots.txt after obtaining the site owner's permission. The bot metadata leaves compliance unconfirmed, and Cloudflare's bot directory marks it as not following robots.txt, so the effective behavior depends partly on the collection configuration.
No general AI search or model-training purpose is documented for this bot. A Funnelback index could be used by the organization that commissioned the crawl, but a visit does not make the page eligible for ChatGPT, Claude, or another unrelated answer engine. Blocking it mainly affects the organization's Funnelback search coverage.
