Promptwatch Logo

Mojeek

Details and information for webmasters regarding Mojeekbot, the web crawler for the Mojeek search engine.
MojeekMojeekBot
Search Engine Crawler

What is Mojeek?

MojeekBot is how Mojeek builds its search index. It fetches pages under the MojeekBot identity, processes index controls, and makes eligible documents available to the Mojeek search engine. A crawl is part of indexing, not a visit from a person who has just clicked a result.

Mojeek says the crawler should make no more than one request per second to a site, whether the request succeeds or fails. It does not support the nonstandard Crawl-delay directive. If observed traffic is faster than that published ceiling, first confirm that all requests are genuine and that logs from multiple hosts have not been combined.

Allowing MojeekBot can make a page eligible for Mojeek's conventional search results. Neither the directory record nor Mojeek's bot page says the crawler collects model-training data. A fetch should not be advertised as increasing visibility in unrelated AI answers.

The crawler reads robots.txt and follows the first record whose user agent contains MojeekBot, falling back to the first wildcard record when no specific section exists. Mojeek also documents support for noindex, nocache, and nofollow meta tags. A noindex page still has to be fetched before the engine can see that instruction.

Relevant for AI search

Is Mojeek relevant for AI search?

Yes. Mojeek collects pages for an AI product, so what it can crawl influences how AI systems describe your brand.

Mojeek feeds a search index that also powers that engine's AI answer features and overviews. One crawl can serve a classic results page and a generated summary, so blocking it costs you both the rankings you would expect and a growing share of AI answers built on the same index.

How to handle Mojeek

Leave public pages crawlable if Mojeek search discovery matters to the site. Use path rules for duplicate or expensive URL spaces, and use a page-level noindex when Mojeek may fetch a document but should not return it in search.

To stop all MojeekBot crawling, publish:

User-agent: MojeekBot
Disallow: /

Mojeek's official page says the crawler honors the Robot Exclusion Standard, though the bot record leaves compliance unset and Cloudflare's bot directory marks it false. Verify the result in access logs. Do not use Crawl-delay to manage this bot because Mojeek says it is unsupported; contact the operator or enforce a rate limit if requests exceed the stated one-per-second ceiling.

Examples

  • A publisher allows MojeekBot to fetch articles but adds `noindex` to a printer-friendly duplicate that should not appear as a separate result.
  • A webmaster verifies a suspicious request with reverse DNS and a matching forward lookup instead of trusting `MojeekBot/0.6` on its own.

Frequently asked questions about Mojeek

Learn about AI visibility monitoring and how Promptwatch helps your brand succeed in AI search.

It crawls web pages for Mojeek's own search index. Eligible pages can then appear in results served by the Mojeek search engine.

Be the brand AI recommends

Monitor your brand's visibility across ChatGPT, Claude, Perplexity, and Gemini. Get actionable insights and create content that gets cited by AI search engines.

Promptwatch Dashboard