Promptwatch Logo

FacebookBot

FacebookBot is a Meta crawler that collects public web content.
MetaFacebookBot
AI CrawlerAI Training

What is FacebookBot?

FacebookBot is a Meta crawler that collects public web content. Meta may use that material to improve language models and other AI products, which is why the bot is categorized for both AI crawling and AI training.

A request means FacebookBot reached a public URL. It does not prove that Meta retained the page, added it to a training dataset, or will reproduce its contents in an AI response. The practical choice for a publisher is whether Meta may fetch the page for this stated development purpose.

FacebookBot is not facebookexternalhit. The latter fetches titles, descriptions, and images when people share links on Facebook, Instagram, or Messenger. Blocking FacebookBot therefore does not express a preference for the link-preview crawler, which needs its own robots.txt group.

It is also separate from the crawlers in Meta's current web crawler documentation. Meta-WebIndexer supports Meta AI search, while Meta-ExternalAgent is described as crawling for model training and product improvement. FacebookBot itself has no documented search-index or citation role, so allowing it should not be presented as a direct way to appear in Meta AI answers.

The stable token in this bot is FacebookBot. No verified IP ranges or Web Bot Auth key directory are attached to the entry, and IP verification is false. A matching user-agent header is useful for log analysis and policy, but another client can copy it.

this bot says FacebookBot respects robots.txt. Sites can allow public material, exclude selected sections, or opt out of the crawl entirely. Policies for Meta's other bots should be written separately because each crawler has its own purpose and token.

Relevant for AI search

Is FacebookBot relevant for AI search?

Yes. FacebookBot collects pages for an AI product, so what it can crawl influences how AI systems describe your brand.

FacebookBot gathers public web content that can end up in the training data for large language models. Once your pages are in that set, they influence how the operator's models talk about you for that model generation. Allowing it lets your own writing carry weight in those answers; blocking it means the models learn about you from third parties instead.

How to handle FacebookBot

Use a FacebookBot-specific rule when the goal is to stop this collection workflow without changing Facebook link previews or the policy for other Meta crawlers. A site-wide opt-out is:

User-agent: FacebookBot
Disallow: /

The bot is recorded as respecting robots.txt. Meta's crawler guidance says robots.txt changes may take up to 24 hours to take effect because the file can be cached, so check later requests before concluding that a new rule was ignored.

Keep private material behind authentication. Since this bot has no authenticated IP range or request-signature method, do not create a security exception from the FacebookBot string alone.

Examples

  • A news site blocks FacebookBot on its article archive but keeps `facebookexternalhit` available so shared links can still receive previews.
  • A software company allows FacebookBot to read public documentation while excluding a licensed research directory with a path-specific rule.
  • A security team labels a request as claimed FacebookBot traffic but refuses to allowlist it because the record provides no verified IP or signature.

Frequently asked questions about FacebookBot

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

Meta operates FacebookBot to crawl public web content that may be used to improve language models and other AI products.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard