Monday, September 14, 2026
TheAI NEWS

The World's Leading Intelligence & Artificial Intelligence Journal

SEO & SearchSep 7, 20264 min read

Site Rules For Applebot-Extended Are Not Considered In Ranking For Search

Apple has updated its official Applebot documentation to explicitly confirm that robots.txt rules for Applebot-Extended have zero impact on search rankings. The change gives publishers full autonomy to block generative AI training without compromising organic visibility in Siri and Spotlight.

Ajinkya Pawar

By Ajinkya Pawar

Head of Search & AI Intelligence • The AI NEWS

Site Rules For Applebot-Extended Are Not Considered In Ranking For Search
Site Rules For Applebot-Extended Are Not Considered In Ranking For Search

Key Developments & Executive Briefing

Executive Briefing
01

Applebot-Extended Excluded from Search Signals

Official DecouplingZero Penalty

Apple officially confirmed that disallowing Applebot-Extended in robots.txt does not impact organic search ranking or indexing in Siri and Spotlight.

02

Granular Opt-Out for Apple Intelligence

Robots.txt Control2 Separate Bots

Publishers can block Apple's foundation model training without losing search traffic, mirroring Google's separation of Googlebot and Googlebot-Extended.

03

Safe Harbor for Paywalled & Proprietary Content

Publisher Autonomy100% Protection

Enterprise media outlets can restrict Apple from ingesting protected editorial archives for generative AI while maintaining full visibility across Apple ecosystem search.

In a crucial clarification for webmasters, enterprise publishers, and technical SEOs worldwide, Apple quietly updated its official Applebot web crawler documentation over the weekend. The update formally decouples web search indexing from generative AI model training, reassuring website operators that restricting Apple's AI scraping will not harm their search rankings.

Ahead of the broad consumer rollout of iOS and Apple Intelligence features across millions of devices, Apple added an unambiguous single-line declaration to its technical documentation:

"Site rules for Applebot-Extended are not considered in ranking for Search."

This brief statement delivers long-awaited legal and technical certainty to digital publishers caught in the tug-of-war between defending proprietary intellectual property and maintaining organic discovery across iOS, iPadOS, macOS, Siri Suggestions, and Spotlight Search.

Origin & Discovery: Documenting the Crawler Split

The quiet documentation change was first spotted and analyzed by Barry Schwartz, News Editor of Search Engine Land and founder of Search Engine Roundtable. Schwartz noted that the move brings Apple directly into alignment with the crawler governance standards pioneered by Google in late 2023.

"Apple has updated its help documentation for Applebot over the weekend and added a single line at the end that reads, 'Site rules for Applebot-Extended are not considered in ranking for Search.' Applebot-Extended is not new; we covered it in April 2025, but here Apple is making it clear, like with Googlebot-Extended, that Applebot-Extended is separate from Search," reported Schwartz. "I guess before the release of iOS to the public, which should be coming out in the next couple of weeks, Apple wanted to make sure the new Siri and Apple Intelligence is clearly defined for what Applebot-Extended can and cannot do with your website."
Applebot Documentation Diff
Applebot Documentation Diff

*Above: Applebot technical documentation diff detailing crawler site rule separations.*

Apple first introduced Applebot-Extended in early 2025 as a dedicated user-agent token designed specifically to train Apple's on-device and server-based Foundation Models. However, until this weekend's documentation revision, technical architects lacked explicit confirmation from Cupertino regarding whether blocking Applebot-Extended would indirectly diminish crawl frequency or algorithmic relevance scores in Siri and Spotlight search indexes.

The Precedent: Replicating the Googlebot-Extended Model

Apple's formal policy codification mirrors the industry-wide crawl governance framework established when Google launched Googlebot-Extended in September 2023. At that time, search engineers demanded that Google separate search retrieval from Vertex AI and Gemini training.

Before crawler separation, publishers faced a severe Hobson's choice: either permit AI web scrapers to harvest paywalled journalism and proprietary research without compensation, or block the search crawler entirely and suffer catastrophic traffic loss.

By creating a strict boundary between Applebot (the core indexing bot powering Safari suggestions, Siri web lookups, and Spotlight) and Applebot-Extended (the data ingestion agent training Apple Intelligence Foundation Models), Apple has closed that dilemma. Webmasters can now selectively deny permission to the AI harvester without sacrificing their presence in Apple search surfaces.

Practitioner Impact: Updating robots.txt Architecture

For enterprise webmasters, digital media networks, and technical SEOs, this confirmation provides immediate actionable direction. Websites that maintain strict IP licensing boundaries or paywalled archives can now update their root robots.txt file with confidence.

To allow Apple's web indexing engine while completely barring Apple's AI models from scraping content for foundation training, technical SEOs should implement the following directive block:

txt
# Allow core Applebot for Spotlight, Safari suggestions, and Siri web search
User-agent: Applebot
Allow: /

# Disallow Applebot-Extended from AI training data ingestion
User-agent: Applebot-Extended
Disallow: /

Key Considerations for SEOs and Media Publishers:

  1. 1.Paywalled Content Protection: Media organizations currently litigating against unauthorized generative AI training—such as the recent lawsuits filed by the Seattle Times and Newsday against OpenAI and Microsoft—can safeguard their content repositories from Apple's pipeline without risking referral drop-offs.
  2. 2.Crawl Budget and Server Load: Disallowing Applebot-Extended reduces unnecessary server overhead from automated scraping runs that do not generate click-through traffic.
  3. 3.Selective Subdirectory Rules: Publishers with hybrid business models can selectively open public marketing pages or press releases to Applebot-Extended while disallowing premium editorial subdirectories (e.g., Disallow: /premium/ or Disallow: /subscriber-only/).

Community Reaction and Industry Consensus

Across search marketing communities and technical discussions on X, reaction to the update has been overwhelmingly positive. Practitioners praised Apple for proactively establishing clear, transparent rules ahead of the imminent public release of Apple Intelligence.

As search engines transform into conversational answer engines and on-device synthesis agents, crawler transparency is becoming a foundational standard. With Apple officially on record, the onus now shifts to publishers to audit their server configurations and align their robots exclusion files with their organization's generative AI data policies.

Discussion (0)

avatar

Be the first to share insights on this story.