---
title: "What is Meta-ExternalAgent?"
description: "What Meta-ExternalAgent (Meta) does with your content, whether to block it, and the exact robots.txt lines — in plain language."
canonical: https://see-geo.com/bots/meta-externalagent
language: en
published: 2026-09-08
updated: 2026-09-08
publisher: SeeGeo
alternates:
  fr: https://see-geo.com/fr/bots/meta-externalagent
  de: https://see-geo.com/de/bots/meta-externalagent
---

# What is Meta-ExternalAgent?

Meta-ExternalAgent is Meta's AI training crawler. Collects pages to train Meta's Llama models.

## What does Meta-ExternalAgent do with your content?

Collects pages to train Meta's Llama models. Like most AI crawlers, Meta-ExternalAgent reads your pages as plain HTML — it does not run JavaScript — so content that only appears after scripts execute is invisible to it.

## Should you block Meta-ExternalAgent?

This is a genuine trade-off. Allowing Meta-ExternalAgent lets future Meta models learn your business exists — useful when customers ask those models for recommendations. Blocking it keeps your content out of training data at the cost of that future visibility. Neither answer is wrong; it depends on what you value more.

## How do you allow or block Meta-ExternalAgent in robots.txt?

Add one of these snippets to the robots.txt file at the root of your site. The first explicitly allows Meta-ExternalAgent everywhere; the second blocks it completely.

```
# Allow Meta-ExternalAgent
User-agent: Meta-ExternalAgent
Allow: /

# Block Meta-ExternalAgent
User-agent: Meta-ExternalAgent
Disallow: /
```

## Is robots.txt enough?

Not always. CDNs and firewalls (Cloudflare in particular) can block crawlers at the network level regardless of what robots.txt says — Cloudflare blocks AI crawlers by default for many accounts. If you want Meta-ExternalAgent to reach your site, check your CDN's bot settings too. SeeGeo's free audit checks both.

## How do I check whether Meta-ExternalAgent can read my site right now?

Two checks, both needed: read your robots.txt for a group naming Meta-ExternalAgent (or the * group it falls back to), then request your homepage with Meta-ExternalAgent's user agent and confirm you get a 200 with your real HTML rather than a 403 or a challenge page. SeeGeo's free crawler check runs both in a few seconds.

[Run the free AI crawler check](https://see-geo.com/ai-crawler-check)

## How does Meta-ExternalAgent compare to similar crawlers?

Meta-ExternalAgent is one of 16 crawlers the SeeGeo audit checks individually. The closest comparisons:

| Crawler | Operator | Role | Runs JavaScript? |
|---|---|---|---|
| Meta-ExternalAgent | Meta | AI training crawler | No |
| GPTBot | OpenAI | AI training crawler | No |
| ClaudeBot | Anthropic | AI training crawler | No |
| Google-Extended | Google | AI training crawler | No |
| CCBot | Common Crawl | AI training crawler | No |

[All 16 crawlers, compared](https://see-geo.com/bots) · [What is an AI crawler?](https://see-geo.com/glossary/ai-crawler)

## Why does crawler access matter? The numbers

Crawler access is where AI visibility starts or ends: a blocked crawler can't read you, and an AI that can't read you can't recommend you.

- Visitors arriving from AI assistants convert at roughly 4.4x the rate of traditional organic search on average (Semrush, cross-industry, 2026).
- AI-assistant referrals are still only about 1% of total web traffic — small, but the fastest-growing acquisition channel measured (multiple 2026 studies).
- Adobe Digital Insights (Q1 2026) measured AI-assistant visitors converting 42% better than non-AI traffic — a full reversal from the year before.
- Cloudflare blocks AI crawlers by default for many accounts, so sites are often invisible to AI without anyone having decided to be.
- Content edits alone — statistics, citations, quotable structure — can raise AI visibility on the order of 30–40% (Princeton GEO study, KDD 2024).

## Frequently asked questions

### What is Meta-ExternalAgent?

Meta-ExternalAgent is Meta's AI training crawler. Collects pages to train Meta's Llama models.

### Should I block Meta-ExternalAgent?

This is a genuine trade-off. Allowing Meta-ExternalAgent lets future Meta models learn your business exists — useful when customers ask those models for recommendations.

### Does Meta-ExternalAgent run JavaScript?

No. Meta-ExternalAgent, like most AI crawlers, reads raw HTML without executing JavaScript. If your content only renders client-side, it is effectively invisible to it.

### Can Meta-ExternalAgent read my website?

Only if two things are true: your robots.txt does not disallow Meta-ExternalAgent, and your server or CDN actually serves the page when Meta-ExternalAgent asks. Firewalls can block it regardless of robots.txt, so the reliable way to know is to fetch your homepage as Meta-ExternalAgent — the free check on see-geo.com/ai-crawler-check does exactly that.

## Related pages

- [What is GPTBot?](https://see-geo.com/bots/gptbot)
- [What is ClaudeBot?](https://see-geo.com/bots/claudebot)
- [What is Google-Extended?](https://see-geo.com/bots/google-extended)
