AI training

Amazonbot

Collects web content for Amazon products and services, potentially including AI training.

See your agent traffic
PurposeAI training
HTTP identifierAmazonbot
IdentityVerify beyond the name

THE SIGNAL

What this visit tells you

Broad collection is not evidence of an Alexa user requesting your page. Separate it from Amazon’s search and user-fetch identities.

Crawl volume
Content paths
Response status

Your next move

Review training permission and crawl load together. Keep business metrics separate from this background collection.

The technical details

Markdown
How to identify it

Look for the Amazonbot identifier in a request’s User-Agent. This is a name match, not identity verification. Version strings may change.

How to check identity

Compare the request IP with the current operator-published ranges linked in the source. A matching name alone does not verify identity.

Access and robots.txt

Use the Amazonbot robots.txt group to control this crawler. Configure Amzn-SearchBot and Amzn-User separately; access does not guarantee inclusion in a product.

Optional full-site opt-out. Merge with existing rules only if intended. robots.txt does not secure private content.

User-agent: Amazonbot
Disallow: /
Seeing 403 or 404 responses?

403 means access was denied; 404 means the resource was not found. Compare the public path, request time, edge security event and origin response to find the cause.

Separate intended restrictions and secret-file probes from pages that should work. A User-Agent name alone does not justify allowing a request.

SourcesReviewed 2026-09-14

Identity and purpose are based on these sources. Analytics interpretation and suggested checks are Apostl guidance.

APOSTL Pulse

See Amazonbot in context

See their requests. Find the pages that matter.

See your agent traffic