Log in
Sign up with Google

AI training crawler · Amazon

Amazonbot is Amazon’s AI training crawler

Amazonbot is Amazon's web crawler used to improve its products and services; its data may be used to train Amazon AI models.

Verifiable operator Respects robots.txt

Our take

Your call

Its data may be used for AI training, so allowing it is mainly a content-licensing decision.

User agent

Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Amazonbot/0.1) Chrome/W.X.Y.Z Safari/537.36

01 · Operator

Who operates Amazonbot

Company
Amazon
Official docs
developer.amazon.com

02 · Behavior

What Amazonbot does

Amazonbot crawls pages automatically to improve Amazon products and services, and the data may be used to train Amazon AI models. It is not triggered by users; Amazon uses Amzn-SearchBot for search and Amzn-User for user requests. It follows robots.txt allow/disallow rules and the noarchive meta tag, but not Crawl-delay.

03 · Impact

Why Amazonbot matters for your site

If you allow it

  • Content may improve Amazon products and services
  • You can opt out of model training per page with the noarchive meta tag

If you block it

  • Data may be used to train Amazon AI models
  • Does not support Crawl-delay to limit load
  • No documented referral traffic

04 · Allow

How to allow Amazonbot

robots.txt

User-agent: Amazonbot
Allow: /

Cloudflare

# Security › WAF › Custom rules › Create rule
Expression: (http.user_agent contains "Amazonbot")
Action:     Skip › All Super Bot Fight Mode rules

# Also check Security › Bots: "Block AI bots" can block it regardless of robots.txt.

WordPress

# WordPress serves a virtual robots.txt. Edit it with your SEO plugin:
# Yoast: SEO › Tools › File editor · Rank Math: General Settings › Edit robots.txt
User-agent: Amazonbot
Allow: /

nginx

# nginx serves every user agent by default.
# Make sure no rule like this blocks it:
# if ($http_user_agent ~* "Amazonbot") { return 403; }

Apache

# Apache serves every user agent by default.
# Make sure .htaccess has no rule like this:
# RewriteCond %{HTTP_USER_AGENT} Amazonbot [NC]
# RewriteRule .* - [F,L]

05 · Block

How to block Amazonbot

Amazonbot follows robots.txt, so one rule is enough. Use a server or CDN rule only to stop spoofed copies.

robots.txt

User-agent: Amazonbot
Disallow: /

Cloudflare

# Security › WAF › Custom rules › Create rule
Expression: (http.user_agent contains "Amazonbot")
Action:     Block

WordPress

# WordPress serves a virtual robots.txt. Edit it with your SEO plugin:
# Yoast: SEO › Tools › File editor · Rank Math: General Settings › Edit robots.txt
User-agent: Amazonbot
Disallow: /

nginx

# In your server { } block:
if ($http_user_agent ~* "Amazonbot") {
    return 403;
}

Apache

# .htaccess
<IfModule mod_rewrite.c>
RewriteEngine On
RewriteCond %{HTTP_USER_AGENT} Amazonbot [NC]
RewriteRule .* - [F,L]
</IfModule>

06 · User agents

User agents we see for Amazonbot

The user agent published by the operator:

Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Amazonbot/0.1) Chrome/W.X.Y.Z Safari/537.36

07 · Verification

Is it really Amazonbot?

Anyone can copy a user agent. Real Amazonbot requests come from the IP ranges published at developer.amazon.com/amazonbot/ip-addresses/.

FAQ

Questions about Amazonbot

Does Amazonbot support Crawl-delay?

No. Amazon states its crawlers honor robots.txt allow and disallow rules but do not support the Crawl-delay directive.

Can I stop Amazon from using my pages for AI training without blocking Amazonbot?

Amazon says its crawlers respect the noarchive robots meta tag, which means do not use the page for model training.

How do I verify Amazonbot?

Match the source IP against the list Amazon publishes at developer.amazon.com/amazonbot/ip-addresses/.

Last reviewed Oct 6, 2026

Your site

See which pages Amazonbot crawls on your website

Log Hero reads your server logs and shows every AI bot, every page, every day.

Sign up with Google

Free during early access