UnsourcedSee who crawls your site →

AI TrainerCohere

cohere-ai: IP ranges, verification & how to handle it

Last verified: 2026-07-16 · maintained by Unsourced

The cohere-ai crawler gathers data for Cohere's enterprise language models. It's a training crawler rather than a live-search one.

What cohere-ai does, and how it differs

The cohere-ai crawler collects text to train Cohere's models, which sell mainly into enterprise and developer products rather than a consumer chat app. Its audience is therefore businesses building on Cohere's API, not end users typing questions, so any visibility it influences is downstream and indirect. It's a training crawler through and through — there's no live-citation path leading back to your site from it.

How to verify cohere-ai

Cohere publishes neither an IP-range feed nor a reverse-DNS footprint for cohere-ai, so there is no way to confirm it by network identity. Treat every request carrying the cohere-ai user-agent as an unverified claim, and judge it on what it does rather than the name it gives.

Should you allow or block cohere-ai?

Recommended: keep. The cohere-ai crawler trains models sold into enterprise products, not a consumer chat app, so any visibility it shapes is indirect — there is no live-citation path from it back to your site.

If you do choose to act in robots.txt (which crawlers honour but don't enforce):

# cohere-ai: recommended to ALLOW — blocking can cost you AI visibility
User-agent: cohere-ai
Disallow:

Official sources

Common questions about cohere-ai

Does the cohere-ai crawler affect consumer AI answers?

Only indirectly. Cohere's models sell into enterprise and developer products rather than a consumer chat app, so any visibility cohere-ai shapes is downstream — there's no live-citation path back to your site.

Is there any referral benefit to allowing cohere-ai?

No. It's a training crawler — it gathers text to train Cohere's models, with no mechanism to cite or link back to the pages it reads.

Can I verify cohere-ai by IP or reverse DNS?

No — Cohere publishes no IP-range feed and no reverse-DNS host for it, so the user-agent can't be confirmed at the network level; treat it as a self-reported claim.

Related crawlers

Are your robots.txt rules for cohere-ai actually holding?

Unsourced audits the requests reaching your site, separates genuine Cohere traffic from spoofed headers, and tracks whether AI answers cite you.

Check your site free →

10-day free trial · no card required · cancel anytime