AgentGrade

Applebot-Extended

Apple · Training crawler · Updated 2026-07-19

Applebot-Extended is Apple’s training crawler — it crawls the web to collect content for model training. Blocking it opts your content out of future models but does not affect live answers.

User-agent string

(none — Applebot-Extended never sends requests; it is a robots.txt token read by Applebot)

Source: vendor-documented at support.apple.com (Applebot page).

Does Applebot-Extended respect robots.txt?

Apple: “Applebot-Extended does not crawl webpages… [it] is only used to determine how to use the data crawled by the Applebot user agent.” Disallowing it opts your content out of training Apple’s foundation models (Apple Intelligence and related features) without affecting Siri, Spotlight, or Safari search.

Allow or block Applebot-Extended

User-agent: Applebot-Extended
Disallow: /

Verifying it’s really Applebot-Extended

Not applicable — no requests carry this identity; verify Applebot instead. User-agent strings can be spoofed by anyone — identity claims are only trustworthy when the source IP matches the operator’s published ranges.

What site owners should know

Apple’s mirror of Google-Extended: a policy switch, not a crawler. Firewall rules for it are meaningless, and pages that disallow it still appear in search results.

Scan your site to see how your robots.txt, content negotiation, and discovery files treat AI agents — including whether you’re blocking agents you meant to allow. Background: the robots.txt guide, how AI agents browse, and agent readiness.

All AI agent user-agents

ClaudeBot · Claude-User · Claude-SearchBot · GPTBot · OAI-SearchBot · ChatGPT-User · PerplexityBot · Perplexity-User · Google-Extended · Google-Agent · Google-CloudVertexBot · Meta-ExternalAgent · Meta-ExternalFetcher · MistralAI-User · Applebot · Applebot-Extended · Amazonbot · bingbot · DuckAssistBot · CCBot · Bytespider