Meta-ExternalAgent

Meta AI training and data collection.

Should you block it?

Your call — blocking is a legitimate choice

Meta-ExternalAgent is Meta's training/data crawler. Blocking it opts your content out of Meta's model training and datasets. For this crawler, training and citation are separate decisions — citations are governed by the operator's retrieval crawler.

Common reasons to block: copyright or licensing policy, competitive sensitivity, paywalled content.
Common reasons to allow: maximum AI reach, research visibility, and being part of the datasets future models learn from.

How to allow it in robots.txt

User-agent: Meta-ExternalAgent
Allow: /

How to block it

User-agent: Meta-ExternalAgent
Disallow: /

Source & verification

Community source Community-maintained list — last reviewed 2026-09-12

https://github.com/ai-robots-txt/ai.robots.txt

We have not verified this entry against the operator's own documentation. Treat the description as indicative, and check the operator's site before relying on it.

"Should you block it?" below is our own interpretation, not the operator's recommendation.

Check your site now

See whether Meta-ExternalAgent — and 39 other AI crawlers — can reach your site today.

Run a free check →

Other Meta crawlers

checkaibots.com — AI visibility checker · Free · No signup