Promptwatch Logo

NewsBank

NewsBank aggregates licensed publisher content for schools, libraries, and government research, learning, and archiving.
NewsBankNewsBank
Archiver

What is NewsBank?

NewsBank works with publishers to build licensed collections of current and archived material. Schools, libraries, government organizations, and other researchers use those collections. Its crawler belongs to that publisher relationship rather than an open project that archives any public website it encounters.

One documented ingestion method is a customizable web harvest of staff-produced content. NewsBank says the setup can exclude syndicated or non-copyrighted material. Publishers can also deliver single-issue PDFs, content-management feeds, and legacy data, so a website crawl is only one way content reaches the archive.

Requests have been recorded as NewsBank.com/1.0 and NewsBank.comMobile/1.0. A participating publisher should compare the requested pages with its agreed harvesting method. NewsBank does not publish a crawler address range or a DNS verification process on the publisher pages, which means the header cannot establish authorization by itself.

The resulting archive is a licensed research product, not a documented generative AI dataset. Blocking an expected harvest can interrupt archive updates or staff access, but allowing it has no stated effect on AI search visibility or model training.

Indirectly relevant

Is NewsBank relevant for AI search?

Indirectly. NewsBank has no AI product of its own, but its output can end up in the systems that AI answers draw on.

NewsBank captures snapshots of pages and stores them, usually for good. Public archives are a common ingredient in AI training datasets, and AI tools sometimes cite an archived version when the live page is gone. What it captures today can still be describing your brand years from now.

How to handle NewsBank

Review the publisher agreement before changing access. Confirm whether NewsBank expects to harvest the website or receive content through another route, since an uncoordinated block can leave a licensed archive out of date.

The known user agents share a stable token, allowing this robots.txt request:

User-agent: NewsBank
Disallow: /

NewsBank does not document robots behavior on its public publisher pages, and the supplied Cloudflare record marks the bot as non-following. Do not treat the rule as guaranteed enforcement. If there is no known NewsBank relationship, ask the operator to identify the collection and its authorization. Use server or edge controls when access must be denied.

Examples

  • A newspaper's configured harvest retrieves locally written stories while leaving syndicated wire articles outside the NewsBank archive.
  • A publisher switches to a CMS feed and confirms the handoff with NewsBank before rejecting the old `NewsBank.com/1.0` requests.

Frequently asked questions about NewsBank

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

The publisher may have licensed material to NewsBank and chosen web harvesting as its delivery method.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard