Google Search / AI search features
Receiver/crawler: Googlebot
Purpose: Search crawling/indexing
Publisher controls: robots.txt; robots meta; canonical; sitemap; structured data
Measurement: Search Console / generative-AI reporting where available
A source-linked matrix of search engines, answer engines, crawlers, retrieval systems, publisher controls and measurement surfaces relevant to Concresca.
Because a search engine is not the internet, and an answer engine is not the only machine receiver. Concresca documents receiver-specific publisher controls from official sources while keeping the underlying content architecture standards-based and platform-independent.
Google, Bing/Copilot, ChatGPT search, Perplexity, Brave, Apple, Claude/Anthropic, Common Crawl, Yandex and IndexNow expose different documented mechanisms. Concresca records those differences rather than pretending one crawler directive has the same meaning everywhere.
Crawler accessibility is implementation readiness, not evidence that a platform has indexed, ranked, cited, trained on or endorsed Concresca.
Bing now documents AI citation and grounding-query measurements. Other platforms expose different or more limited publisher measurements. Concresca will label unavailable live metrics NOT CONNECTED rather than zero.
The matrix records documented publisher controls and measurement surfaces. It does not claim that any receiver has crawled, indexed, ranked, cited, trained on, or endorsed Concresca.
Receiver/crawler: Googlebot
Purpose: Search crawling/indexing
Publisher controls: robots.txt; robots meta; canonical; sitemap; structured data
Measurement: Search Console / generative-AI reporting where available
Receiver/crawler: bingbot
Purpose: Search crawling and AI grounding
Publisher controls: robots.txt; sitemap; IndexNow
Measurement: Bing Webmaster Tools AI Performance
Receiver/crawler: OAI-SearchBot
Purpose: Search discovery / answer retrieval
Publisher controls: robots.txt
Measurement: Referral/server logs; official publisher controls; no performance claim without measurement
Receiver/crawler: GPTBot
Purpose: Model development/training access control
Publisher controls: robots.txt
Measurement: Crawler logs where available
Receiver/crawler: PerplexityBot
Purpose: Search indexing/retrieval
Publisher controls: robots.txt
Measurement: Referral/server logs where available
Receiver/crawler: Brave search crawler
Purpose: Independent search crawling
Publisher controls: robots.txt / documented Brave controls
Measurement: Brave search visibility where measurable
Receiver/crawler: Applebot
Purpose: Search/discovery and contextual grounding
Publisher controls: robots.txt
Measurement: Server logs where available
Receiver/crawler: Applebot-Extended
Purpose: Publisher control for model use of Applebot-crawled content
Publisher controls: robots.txt
Measurement: Crawler/control audit
Receiver/crawler: Claude-SearchBot; Claude-User
Purpose: Search and user-directed retrieval
Publisher controls: robots.txt
Measurement: Server logs where available
Receiver/crawler: ClaudeBot
Purpose: Model development crawling
Publisher controls: robots.txt
Measurement: Crawler logs where available
Receiver/crawler: CCBot
Purpose: Open web corpus crawling
Publisher controls: robots.txt; sitemaps
Measurement: Common Crawl index/server logs where available
Receiver/crawler: YandexBot
Purpose: Search crawling/indexing
Publisher controls: robots.txt
Measurement: Yandex Webmaster where available
Receiver/crawler: n/a
Purpose: URL-change notification, not crawling itself
Publisher controls: IndexNow key + submission API
Measurement: Submission logs / observed recrawl only after authorized live submission
Source links establish traceability and support. They do not imply that the source endorses Concresca’s constitutional proposals.
Documents Google robots.txt interpretation and crawler access rules.
Google Search Central — official/source page ↗Documents AI citation counts, cited pages, grounding queries, and citation trends.
Microsoft Bing — official/source page ↗Documents OAI-SearchBot controls for ChatGPT search and distinguishes search discovery from GPTBot model-training controls.
OpenAI — official/source page ↗Documents PerplexityBot robots.txt behavior for search indexing.
Perplexity — official/source page ↗Documents Brave Search crawling behavior and publisher controls.
Brave — official/source page ↗Documents Applebot robots controls and Applebot-Extended publisher controls for foundation-model use.
Apple — official/source page ↗Documents ClaudeBot, Claude-SearchBot, and Claude-User as separate crawlers with separate controls.
Anthropic — official/source page ↗Documents CCBot, robots.txt, sitemaps, and crawl inclusion limits.
Common Crawl — official/source page ↗Documents Yandex robots.txt behavior.
Yandex — official/source page ↗Documents URL-change notification and ownership verification.
IndexNow — official/source page ↗Documents participating search engines and notification sharing.
IndexNow — official/source page ↗NO JUDGMENT WHATSOEVER. Concresca coordinates without assigning moral worth, character, guilt, danger, trustworthiness, loyalty, purity, normality, or social standing. Questions, thoughts, identities, messages, content, and conduct are not objects of Concresca judgment.