Skip to content
MarketHQ
AI crawler

Amazonbot

Amazon's ai model training crawler.

Quick answer

Amazonbot is Amazon's general-purpose web crawler. It gathers publicly available content from across the web that Amazon uses to improve the accuracy of its products and services, and that same content may also be used to train Amazon's own AI models, separate from the crawlers other companies run for their assistants.

What it does

Amazonbot crawls public web pages so Amazon can improve the accuracy of its products and services, and content it collects may also be used to train Amazon's AI models. It is a general-purpose data-gathering crawler for Amazon broadly, not a bot tied to one specific feature like search results or shopping listings. Amazonbot honors the Robots Exclusion Protocol, meaning it reads and follows the user-agent and allow or disallow directives set in a site's robots.txt file. Amazon publishes a dedicated IP address list for Amazonbot on its developer site, so site owners can confirm that traffic claiming to be Amazonbot is genuine by checking it against that published list.

Facts

Operator

Amazon

Robots token

Amazonbot

User agent string(s)

  • Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Amazonbot/0.1) Chrome/W.X.Y.Z Safari/537.36

Purpose

AI model training

Respects robots.txt

Yes

IP ranges

https://developer.amazon.com/amazonbot/ip-addresses/

From Amazon's documentation, checked Oct 2026.

Should you block it?

Disallowing Amazonbot keeps a site's future content out of the pool Amazon uses to improve its products and train its AI models, which matters to publishers who want tighter control over how their work gets reused. The tradeoff is less about visibility in a well known AI assistant's answers and more about whether Amazon's broader product ecosystem, which includes shopping and voice features many consumers use, has accurate, up to date information about a brand. A company whose customers research products on Amazon-adjacent surfaces may want its public pages crawled so that ecosystem reflects it correctly, while a publisher mainly worried about content reuse for AI training may prefer to disallow it. Amazonbot respects standard robots.txt rules, so a disallow rule takes effect without any extra verification step.

How to block it

Add this to robots.txt to block Amazonbot: User-agent: Amazonbot Disallow: / Amazonbot respects the Robots Exclusion Protocol, so this rule stops it from crawling once it is in place.

Frequently asked questions

What does Amazonbot do with the content it crawls?

Amazon says it uses crawled content to improve its products and services, and that content may also be used to train Amazon AI models.

Does Amazonbot follow robots.txt?

Yes. Amazon says its crawlers, including Amazonbot, respect the Robots Exclusion Protocol and honor user-agent and allow or disallow directives.

Is Amazonbot the same bot that powers Alexa's live answers?

No. Amazonbot is a proactive crawler for training and product improvement, separate from Amzn-User, which fetches pages for a specific user request.

Related

  • AI crawler directory — every AI crawler's robots.txt token, purpose and facts in one place.
  • Free AI crawler checker — check which AI crawlers a domain's robots.txt actually allows or blocks.
  • Glossary — definitions of the AI-visibility terms that come up alongside crawler behavior.

Sources

  • Amazonbot is used to improve Amazon's products and services and may be used to train Amazon AI models. Source: https://developer.amazon.com/en/amazonbot (checked Oct 2026)
  • Amazon's bots respect the Robots Exclusion Protocol, honoring the user-agent and allow/disallow directives. Source: https://developer.amazon.com/en/amazonbot (checked Oct 2026)

MarketHQ tracks brand mentions across communities, news, blogs, social and AI answers, and turns them into gap analysis and action plans.