resulthack

All AI crawlers /

Cotoyogi

ROIS-DS · Unconfirmed · robots.txt: Yes

A research crawler run by the Center for Research and Development on Data Lake at ROIS-DS, collecting Japanese language data resources from the web.

Block Cotoyogi

User-agent: Cotoyogi
Disallow: /

Unconfirmed by the operator ROIS-DS's documentation we checked does not say what blocking Cotoyogi changes beyond the crawler itself.

Allow Cotoyogi

User-agent: Cotoyogi
Allow: /

An allow group only changes anything if a broader rule would block Cotoyogi. Under the robots.txt standard (RFC 9309, section 2.2.1) a crawler follows the group that names it and uses the User-agent: * group only when no group does — so this group lets it in even if your * group says Disallow: /. Token matching is case-insensitive.

Put these lines in /robots.txt at the root of each host (each subdomain has its own file). robots.txt is a request to well-behaved crawlers, not access control.

Who blocks it

Cotoyogi is not in the 2026-10-09 survey: it was added to this directory after we last read the most-visited sites' robots.txt files, so we have no count for it yet.

Facts

Operator
ROIS-DS
Purpose
Unconfirmed by the operator

ROIS-DS documents it as a crawler run by its Center for Research and Development on Data Lake for collecting Japanese language data resources. It does not say what the collected data is used for.

Obeys robots.txt
Yes. ROIS-DS gives robots.txt (RFC 9309) as the first way to stop Cotoyogi, with Disallow (including * and $ patterns) and Crawl-delay examples, and says a robots meta nofollow tag stops it following a page’s links.
In your server logs
The user-agent string contains "(compatible; Cotoyogi/4.0;".
How to verify it
ROIS-DS publishes the range of addresses it crawls from on its crawler page: 157.1.136.4 to 157.1.136.11.

What ROIS-DS says

Quoted word for word from the operator's documentation.

“Cotoyogi is a Web crawler (aka robot) operated at Center for Research and Development on Data Lake, ROIS-DS for collecting Japanese language data resources.”

“For example, the following forbids Cotoyogi to retrieve any content from your site.”

“If the repetitive accesses are annoying to you, please follow the Robots Exclusion methods or contact us as described below.”

“For example, the following directs Cotoyogi to access the site at most once per 30 seconds.”

“In a nutshell, if you put <META NAME="robots" CONTENT="nofollow"> in the HTML headers, Cotoyogi will not follow the links found in the documents.”

“For more details, please refer to RFC 9309.”

“Disallow accepts wildcard character " * " and end-of-path designator " $ " as well.”

“User agent string: Mozilla/5.0 (compatible; Cotoyogi/4.0;”

“Range of IP addresses: 157.1.136.4 - 157.1.136.11”

Sources

Same purpose, other operators

← All 49 AI crawlers