Cohere legacy training-data crawler
Your call — blocking is a legitimate choice
cohere-training-data-crawler is Cohere's training/data crawler. Blocking it opts your content out of Cohere's model training and datasets. It does not remove you from AI answers — citations are governed by separate retrieval crawlers.
Common reasons to block: copyright or licensing policy, competitive sensitivity, paywalled content.
Common reasons to allow: maximum AI reach, research visibility, and being part of the datasets future models learn from.
User-agent: cohere-training-data-crawler Allow: /
User-agent: cohere-training-data-crawler Disallow: /
See whether cohere-training-data-crawler — and 39 other AI crawlers — can reach your site today.