Amazonbot: a search crawler run by Amazon
Amazonbot is a search crawler. It reads pages for the index behind this product: Alexa and Rufus. Operator: Amazon.
At a glance
- Operator
- Amazon
- robots.txt token
Amazonbot- Purpose
- Search. Search crawlers read pages to build the index an assistant searches when it answers. A page they cannot read cannot be found or cited in those answers.
- What it feeds
- Alexa and Rufus
What blocking it changes
If you keep Amazonbot out, your pages cannot be found and cited in the search answers of this product: Alexa and Rufus.
User agent
The full user agent, from the operator’s documentation. A request from Amazonbot carries this text, and a server log shows it.
Mozilla/5.0 AppleWebKit/537.36 (KHTML, like Gecko; compatible; Amazonbot/0.1; +https://developer.amazon.com/support/amazonbot) Chrome/119.0.6045.214 Safari/537.36robots.txt lines for Amazonbot
robots.txt is a plain text file at the root of a site (yoursite.com/robots.txt). A group starts with one or more User-agent lines that name crawlers, followed by Allow and Disallow lines for paths. A crawler named in a group follows the groups that name it and ignores the group for every crawler (User-agent: *).
To let it in
User-agent: Amazonbot
Allow: /To keep it out
User-agent: Amazonbot
Disallow: /Does your site let it in?
The free crawler check reads your robots.txt for Amazonbot and requests your page with its user agent, so it shows whether a robots.txt rule or a firewall turns it away.