Promptwatch Logo

CitibotSiteCrawler

CitibotSiteCrawler collects public data from government websites to power Citibot’s AI civic engagement tools.
CitibotCitibotSiteCrawler
AI Crawler

What is CitibotSiteCrawler?

CitibotSiteCrawler is Citibot's crawler for public information on government websites. The collected material supports Citibot's civic engagement tools, which government agencies use to answer resident questions through their own service channels.

Citibot has not published a robots.txt commitment for this identity. Cloudflare's bot directory marks the crawler as not following robots.txt, while the bot's listing leaves the status unknown. A robots rule can state the site's preference, but an agency that must prevent access should use a server, firewall, or authentication control as well.

Citibot describes its AI as a closed system that retrieves information from an agency's approved knowledge base rather than searching the open web for every answer. In that setup, the crawler supplies source material from public government pages so the deployed service can use official content.

Citibot does not publish a crawler page that explains visit schedules, page depth, or refresh rules. The directory records one user agent, CitibotSiteCrawler/1.0. Logs can show which government URLs it fetched, but the available operator material does not support assumptions about when it will return.

There is no Web Bot Auth signature URL or verified IP status attached to the record. CitibotSiteCrawler is useful for classifying a request, but any client can copy that header. Confirm unexpected traffic with Citibot before giving it access that would otherwise be denied.

These crawls can affect the information available in a participating agency's Citibot service. They are not documented as general web indexing or foundation-model training. A CitibotSiteCrawler visit therefore says nothing by itself about rankings in conventional search or inclusion in a public AI answer engine.

Relevant for AI search

Is CitibotSiteCrawler relevant for AI search?

Yes. CitibotSiteCrawler collects pages for an AI product, so what it can crawl influences how AI systems describe your brand.

CitibotSiteCrawler fetches, indexes, or retrieves pages for an AI product, so the pages it can reach shape how that system describes your brand and products.

How to handle CitibotSiteCrawler

A government agency using Citibot should agree with the operator on the public pages that belong in its approved source set. Keep drafts, staff systems, and any resident data outside the crawler's reachable surface. A non-customer has no documented search or training reason to allow the bot.

Publish this rule if the site does not consent to crawling:

User-agent: CitibotSiteCrawler
Disallow: /

Robots.txt is not a reliable enforcement point here because Citibot makes no public commitment and the Cloudflare directory flag is false. Watch for continued requests, then block the token at the edge if necessary. Do not create an allow rule for protected content merely because the request presents the expected user agent.

Examples

  • A city deploying Citibot reviews the crawler's URL list and limits it to published service pages that staff have approved for resident answers.
  • A government agency with no Citibot deployment sees CitibotSiteCrawler in its logs, asks the operator to confirm the activity, and enforces a block when no crawl is expected.

Frequently asked questions about CitibotSiteCrawler

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

It collects public data from government websites for Citibot's AI civic engagement tools.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard