GEO Glossary

Meta-ExternalAgent

Meta-ExternalAgent is Meta's crawler for training AI models and indexing content. How it differs from Meta-ExternalFetcher and Meta-WebIndexer, its user agent string, and how to control it in robots.txt.

By Ramanath, CTO & Co-Founder at Presenc AI · Last updated: October 1, 2026

What Is Meta-ExternalAgent?

Meta-ExternalAgent is the web crawler Meta uses to collect content for its AI systems. Meta's developer documentation says it "crawls the web for use cases such as training foundation AI models or improving products by indexing content directly." It identifies itself as meta-externalagent/1.1, and the robots.txt token is meta-externalagent.

Meta's AI Crawlers Compared

Meta documents three AI-related agents plus the older link-preview crawler. They do different jobs and follow robots.txt differently.

AgentPurpose, in Meta's wordsrobots.txt
meta-externalagentTraining foundation AI models or improving products by indexing content directlyStandard rules apply
meta-externalfetcherFetches individual links at a user's request and supports evaluating and improving agentic AI capabilitiesMeta says this crawler may bypass robots.txt rules
meta-webindexerNavigates the web to improve Meta AI search result qualityControlled by its own token
facebookexternalhitCrawls content shared on Meta's apps to build link previewsMay bypass rules for security or integrity checks

Meta says robots.txt changes can take up to 24 hours to register.

robots.txt Examples

To opt out of training while staying eligible for Meta AI search results:

User-agent: meta-externalagent
Disallow: /

User-agent: meta-webindexer
Allow: /

Blocking meta-externalagent does not stop meta-externalfetcher, which acts on a user's request and which Meta says may bypass robots.txt. If you need to refuse it, that has to happen at the firewall or CDN.

Why It Matters for AI Visibility

Meta AI is built into WhatsApp, Instagram, Facebook, and Meta's glasses, and Meta now runs a separate agent, Muse, that browses and buys on a user's behalf. What these products know about a brand comes partly from what Meta's crawlers could read. A site that blocks all Meta agents is opting out of that entire surface. A site that blocks only the training crawler keeps search and user-initiated fetches working. See Meta AI citation patterns and Meta Muse: what it means for brands.

Commonly Confused With

Meta-ExternalAgent is not facebookexternalhit, the link-preview crawler that has existed for years and has nothing to do with AI training. It is also not the browser a Muse agent uses to shop on a site, which Amazon has said does not identify itself. That dispute is covered in Amazon blocks Meta Muse.

Frequently Asked Questions

Meta documents two forms: meta-externalagent/1.1 followed by a link to its crawler documentation, or the short form meta-externalagent/1.1. The robots.txt token is meta-externalagent.
Meta-ExternalAgent crawls the web for AI training and indexing. Meta-ExternalFetcher fetches a single link when a user asks for it, and Meta says it may bypass robots.txt rules because the request is user-initiated.
Meta's documentation lists meta-externalagent as a robots.txt token and says changes can take up to 24 hours to register. It reserves a bypass only for meta-externalfetcher and for security checks by facebookexternalhit.
Block it if you do not want your content used to train Meta's models. If you still want to appear in Meta AI answers, leave meta-webindexer allowed, since that agent supports Meta AI search results.

Track Your AI Visibility

See how your brand appears across ChatGPT, Claude, Perplexity, and other AI platforms. Start monitoring today.