Log in
Sign up with Google

Search engine crawler · TutuFind (图图搜索)

TutusooBot is TutuFind (图图搜索)’s search engine crawler

TutusooBot is the crawler of TutuFind (图图搜索), a Chinese search service. It builds an index of public web pages.

Respects robots.txt

Our take

Allow

It is a search engine crawler that follows robots.txt and can list your pages in search.

User agent

TutusooBot/0.1 (+https://tutufind.com/bot)

Activity

How TutusooBot crawls

Requests per site per dayLAST 8 DAYS
Sep 30Oct 2Oct 5Oct 7
What it crawls30 DAYS
  • Content pages 55%
  • Sitemaps & feeds 39%
  • Home page 3.4%
  • Not found (404) 2.3%

01 · Operator

Who operates TutusooBot

Official docs
tutufind.com

02 · Behavior

What TutusooBot does

TutusooBot crawls public web pages for the TutuFind index. It follows robots.txt by default and avoids private network addresses. It waits 1 second between page requests to a site by default and honors Crawl-delay up to a cap. Pages disallowed later are removed from the index on recrawl.

03 · Impact

Why TutusooBot matters for your site

If you allow it

  • Your pages can appear in TutuFind search
  • It follows robots.txt and Crawl-delay

If you block it

  • TutuFind is a small search engine with little referral traffic for most sites
  • No IP ranges are published for verification

04 · Allow

How to allow TutusooBot

robots.txt

User-agent: TutusooBot
Allow: /

Cloudflare

# Security › WAF › Custom rules › Create rule
Expression: (http.user_agent contains "TutusooBot")
Action:     Skip › All Super Bot Fight Mode rules

# Also check Security › Bots: "Block AI bots" can block it regardless of robots.txt.

WordPress

# WordPress serves a virtual robots.txt. Edit it with your SEO plugin:
# Yoast: SEO › Tools › File editor · Rank Math: General Settings › Edit robots.txt
User-agent: TutusooBot
Allow: /

nginx

# nginx serves every user agent by default.
# Make sure no rule like this blocks it:
# if ($http_user_agent ~* "TutusooBot") { return 403; }

Apache

# Apache serves every user agent by default.
# Make sure .htaccess has no rule like this:
# RewriteCond %{HTTP_USER_AGENT} TutusooBot [NC]
# RewriteRule .* - [F,L]

05 · Block

How to block TutusooBot

TutusooBot follows robots.txt, so one rule is enough. Use a server or CDN rule only to stop spoofed copies.

robots.txt

User-agent: TutusooBot
Disallow: /

Cloudflare

# Security › WAF › Custom rules › Create rule
Expression: (http.user_agent contains "TutusooBot")
Action:     Block

WordPress

# WordPress serves a virtual robots.txt. Edit it with your SEO plugin:
# Yoast: SEO › Tools › File editor · Rank Math: General Settings › Edit robots.txt
User-agent: TutusooBot
Disallow: /

nginx

# In your server { } block:
if ($http_user_agent ~* "TutusooBot") {
    return 403;
}

Apache

# .htaccess
<IfModule mod_rewrite.c>
RewriteEngine On
RewriteCond %{HTTP_USER_AGENT} TutusooBot [NC]
RewriteRule .* - [F,L]
</IfModule>

06 · User agents

User agents we see for TutusooBot

User agentShareLast seenStatus
TutusooBot/0.1 (+https://tutufind.com/bot) 100% Unverified

07 · Verification

Is it really TutusooBot?

TutusooBot’s operator publishes no IP ranges or hostnames, so requests cannot be verified. Treat the user agent as a claim, and watch your logs for unusual request rates.

FAQ

Questions about TutusooBot

What is TutusooBot?

TutusooBot is the web crawler of TutuFind (图图搜索). It indexes public web pages for that search service.

How do I block TutusooBot?

Add User-agent: TutusooBot with Disallow: / to robots.txt. Previously indexed pages that become disallowed are removed on recrawl.

Does TutusooBot honor Crawl-delay?

Yes. TutuFind says Crawl-delay is honored, with a reasonable upper limit. Without it, the default interval is 1 second.

Last reviewed Oct 8, 2026

Your site

See which pages TutusooBot crawls on your website

Log Hero reads your server logs and shows every AI bot, every page, every day.

Sign up with Google

Free during early access