facebook.com

Classified by Aater · 28 Jun 2026

Primary bottleneck: Reachability

Reachability
blocked
Legibility
not evaluated
Content Type
not evaluated
Structure is the baseline — what AI agents actually do is the full picture
Structure
Reachability
Legibility
Content Type
Observe
Which AI agents visit facebook.com?
Install Pulse →
Understand
Do Claude & ChatGPT cite you?
Install Pulse →
Govern
Blocks ChatGPT, Claude, Gemini, Perplexity

Declared in robots.txt

Structure tells you whether AI agents can reach and read facebook.com's content. Whether they actually do — visiting, citing, following your rules — is what Observe and Understand show, via Pulse.

How this classification was derived · the three gates
Gate 1 · Reachability
blocked
⚠ Primary bottleneck

All crawlers are blocked via robots.txt. AI systems cannot access this content.

Gate 2 · Legibility
Not evaluated

Awaiting Gate 1 — legibility is not evaluated when the domain is unreachable.

Gate 3 · Content Type
Not evaluated

Awaiting Gate 1 — content type is not evaluated when the domain is unreachable.

Agent participation · crawler access (robots.txt)
ChatGPT
Blocked
Claude
Blocked
Gemini
Blocked
Perplexity
Blocked

Keyed to each system's inference fetcher — the agent that fetches at answer-time. Blocking a training crawler (e.g. GPTBot) does not by itself block citation. Does not affect the structural gate result.

Primary bottleneck: Reachability

Reachability is the gate below threshold.

AI systems can't reliably reach facebook.com's content.

Recommended next steps
1.Add specific verifiable claims to key pages

Expected impact: Restores crawler access

2.Add authorship markup

Expected impact: Restores crawler access

View implementation →
<script type="application/ld+json">
{
  "@context": "https://schema.org",
  "@type": "Article",
  "author": { "@type": "Person", "name": "Author Name" },
  "datePublished": "2024-01-15"
}
</script>
3.Add publication dates

Expected impact: Restores crawler access

View implementation →
<time datetime="2024-01-15" pubdate>Published January 15, 2024</time>
Why these recommendations?

Robots.txt is blocking the answer-time fetchers that produce AI citations.

Renders as an empty JavaScript shell — a non-rendering AI crawler receives almost no content.

No structured data (JSON-LD / schema.org) — machine-readable metadata is absent.

No <main>/<article> landmark — main content is not demarcated for parsers.

No <h1> heading — weak document structure.

No Organization schema — entity identity is not machine-asserted.

No author attribution — content lacks attributable provenance.

Add specific verifiable claims to key pages: AI systems extract attributable facts — named figures, dated events, quantities. Generic statements do not register.

Add authorship markup: AI systems weight content from named authors more heavily. Anonymous content is treated as lower trust.

Add publication dates: Retrieval-augmented AI pipelines filter by recency. Undated content is deprioritised in freshness-weighted retrieval.

Entity corroboration
Observational · does not affect the gate result

Can this entity be independently verified outside its own domain? Each source the public web recognizes makes the entity easier for an AI system to trust. Observational only — a lower bound: absence means “not documented in the knowledge graph,” not “does not exist.”

Recognized as an entity in Wikidata: Facebook.
English Wikipedia article present.
Linked YouTube presence.
·No linked GitHub presence.
Crunchbase organization profile present.
See how facebook.com compares to a competitor →
Pulse · Live monitoring
This is a snapshot. Pulse is the live feed.

See which AI agents are actually crawling facebook.com, and separate the training crawlers (GPTBot, ClaudeBot, Google-Extended) from the citation-path fetchers(ChatGPT-User, Claude-User, Perplexity-User) that fetch at answer-time. A lightweight snippet reveals the agent activity your logs don't surface.

Unknown detected → JS Lite
JS-only on this stack — answer-time agents only

When the Reachability gate is resolved, Pulse will surface real-time AI agent activity on facebook.com — separating training crawlers from citation-path fetchers.

Activate Pulse on facebook.com

Lightweight install · no performance impact · first agent activity in 24–72h (traffic-dependent)

Measured 28 Jun 2026 · 22d ago

This classification reflects Aater's assessment of observable structural signals. It does not represent an editorial opinion about the quality or value of this domain or organisation. Domain owners may request removal by writing to founder@aater.ai. Requests are honoured within 48 hours.