resulthack

All AI crawlers /

bedrockbot

Amazon · Unconfirmed · robots.txt: Yes

Crawls the URLs an AWS customer selects for an Amazon Bedrock knowledge base, and on each run fetches every page reachable from them within the customer's scope and filters.

Block bedrockbot

User-agent: bedrockbot
Disallow: /

Unconfirmed by the operator Amazon's documentation we checked does not say what blocking bedrockbot changes beyond the crawler itself.

Allow bedrockbot

User-agent: bedrockbot
Allow: /

An allow group only changes anything if a broader rule would block bedrockbot. Under the robots.txt standard (RFC 9309, section 2.2.1) a crawler follows the group that names it and uses the User-agent: * group only when no group does — so this group lets it in even if your * group says Disallow: /. Token matching is case-insensitive.

Put these lines in /robots.txt at the root of each host (each subdomain has its own file). robots.txt is a request to well-behaved crawlers, not access control.

Who blocks it

bedrockbot is not in the 2026-10-09 survey: it was added to this directory after we last read the most-visited sites' robots.txt files, so we have no count for it yet.

Facts

Operator
Amazon
Purpose
Unconfirmed by the operator

Amazon documents it as the Web Crawler of Amazon Bedrock knowledge bases: an AWS customer selects URLs and the crawler fetches them into that customer's knowledge base. That is none of training, a public AI search index or a user fetch as Amazon states it.

Obeys robots.txt
Yes. Amazon says the Bedrock Web Crawler respects robots.txt in accordance with RFC 9309. It looks first for rules naming bedrockbot-UUID and then for generic bedrockbot rules.
In your server logs
Unconfirmed by the operator Amazon's documentation we checked does not give its user-agent string.
How to verify it
Unconfirmed by the operator Amazon's documentation we checked gives no IP list or DNS check for this crawler, so a request claiming to be bedrockbot cannot be verified against the operator.

What Amazon says

Quoted word for word from the operator's documentation.

“The crawler will first look for bedrockbot-UUID rules and then for generic bedrockbot rules in the robots.txt file.”

“The Web Crawler respects robots.txt in accordance with the RFC 9309.”

“The Amazon Bedrock provided Web Crawler connects to and crawls URLs you have selected for use in your Amazon Bedrock knowledge base.”

“Each time the the Web Crawler runs, it retrieves content for all URLs that are reachable from the source URLs and which match the scope and filters.”

Sources

Other crawlers from Amazon

Same purpose, other operators

← All 49 AI crawlers