Promptwatch Logo

MRGbot

Search engine aimed at generating a corpus of data to be able to aggregate data in various ways.
MRG Web Services srlMRGbot
Search Engine Crawler

What is MRGbot?

MRG Web Services srl describes MRGbot as a crawler for building a data corpus that can be aggregated in different ways. The record classifies it as a search engine crawler, but it does not name a public search interface or explain which aggregations are produced from that corpus.

Recorded requests contain MRGbot/1.0. One full example links to https://www.mrg.ro/bot.html, while another stored example is cut off after the product token. Matching MRGbot is therefore more reliable than requiring a complete user-agent string.

A general data corpus could support many downstream uses, but the record does not identify language-model training or generated search answers as one of them. Allowing MRGbot may contribute pages to MRG's own aggregation system; it has no documented effect on mainstream search rankings or AI citations.

The practical workflow is a conventional fetch-and-index cycle: the bot requests public URLs, receives the site's response, and contributes what it can process to MRG's corpus. The supplied evidence does not document URL discovery, revisit frequency, content retention, or a fixed source network. Those questions cannot be answered from the name alone.

Relevant for AI search

Is MRGbot relevant for AI search?

Yes. MRGbot collects pages for an AI product, so what it can crawl influences how AI systems describe your brand.

MRGbot feeds a search index that also powers that engine's AI answer features and overviews. One crawl can serve a classic results page and a generated summary, so blocking it costs you both the rankings you would expect and a growing share of AI answers built on the same index.

How to handle MRGbot

Allow the crawler only if contributing public pages to MRG's corpus fits the site's distribution policy. Since the downstream product is not described in detail, organizations with licensing or reuse restrictions may prefer a narrow allow list or a block while they seek clarification.

To opt out across the site, use:

User-agent: MRGbot
Disallow: /

The MRGbot metadata and Cloudflare directory both mark this corpus crawler as following robots.txt. Confirm that with request logs after changing the file. No IP list or signed-agent directory is supplied, so do not create a broad network allow rule from the user agent alone.

Examples

  • A log pipeline matches `MRGbot/1.0` even when the rest of the user-agent value is incomplete, avoiding a false count of ordinary browsers.
  • A webmaster adds a full disallow and checks later requests for 200 responses to determine whether the crawler applied the new policy.

Frequently asked questions about MRGbot

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

The directory attributes MRGbot to MRG Web Services srl.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard