cohere-ai user agent

Operated by Cohere · Model training · Learning from you

robots.txt token
cohere-ai

cohere-ai identifies itself with the token cohere-ai. That is the case-insensitive substring to grep for in access logs, and the exact name to put after User-agent: in robots.txt. It collects pages as training data for a model.

Allow or block cohere-ai in robots.txt

Robots.txt applies the most specific matching group, so a group naming cohere-ai overrides your User-agent: * rules for this bot alone.

Allow
User-agent: cohere-ai
Allow: /
Block
User-agent: cohere-ai
Disallow: /

What blocking costs you: Cohere's enterprise models won't train on your content.

To see which AI crawlers your robots.txt allows right now, run your domain through the AI crawler access checker. To check whether AI answers actually cite you, use the AI citation checker.

FAQ

What is the cohere-ai user agent?
cohere-ai announces itself with the User-Agent token "cohere-ai". That token is the case-insensitive substring to match in server logs and the exact name to put after "User-agent:" in robots.txt. Cohere does not publish one fixed full string, so match on the token rather than the whole line.
Who operates cohere-ai and what does it collect?
cohere-ai is run by Cohere and collects pages as training data for a model. In the AI picture that puts it under "Learning from you": Feeding the models that power future answers.
How do I block cohere-ai?
Add a robots.txt group naming the token exactly: User-agent: cohere-ai followed by Disallow: /. Robots.txt matches the most specific group, so a cohere-ai group overrides whatever your User-agent: * group says. Cohere's enterprise models won't train on your content.
Should I block cohere-ai?
That depends on what you lose. Cohere's enterprise models won't train on your content. Sites chasing AI citations usually keep the live-answer and AI-search bots open and make a separate decision about the training crawlers. Run your domain through the AI crawler access checker to see which ones your robots.txt lets in today.
Can cohere-ai be spoofed?
Yes. A User-Agent header is self-reported, so any scraper can claim to be cohere-ai. The token is fine for reporting and for robots.txt, but if you are rate-limiting or firewalling on it, verify the request IP against Cohere's published ranges before you trust the name.

Paste a full User-Agent line and identify any crawler with the AI bot user agent list and lookup, or see every token side by side.