Two Webs

Top 100 / digicert.com

65

Two different webs

2 of 14 AI agents get the full page at https://digicert.com/. Checked 2026-09-25 09:17 UTC.

digicert.com scores 65 out of 100 for machine visibility, number 68 of the 100 most visited websites checked on 2026-09-25. 2 of 14 AI agents receive the full page: 2 of 8 training crawlers, 0 of 3 AI search indexes and 0 of 3 assistant fetchers. Googlebot itself is partial: Page allowed but marked noindex. The site publishes an llms.txt.

youfull pagedegradedblockedunknown

Human

Status
200 → https://www.digicert.com/
Title
Pardon Our Interruption
H1
Pardon Our Interruption
Words
111
Scripts
5 tags · 3 KB
Meta robots
noindex, nofollow
Canonical
none
Server
hidden

Machine

robots.txt
23 groups · 15 sitemaps
llms.txt
present
llms-full.txt
missing
X-Robots-Tag
none
Blocked
none
Degraded
GPTBot, OAI-SearchBot, ChatGPT-User, ClaudeBot, Claude-SearchBot, Claude-User, PerplexityBot, Perplexity-User, Meta-ExternalAgent, Amazonbot, Bytespider, CCBot
Challenged
none

Gaps

  • medium
    14 agents get a degraded pageGooglebot: Page allowed but marked noindex · GPTBot: Page allowed but marked noindex · OAI-SearchBot: Different content than humans receive · ChatGPT-User: Different content than humans receive · ClaudeBot: Page allowed but marked noindex · Claude-SearchBot: Page allowed but marked noindex · Claude-User: Page allowed but marked noindex · PerplexityBot: Page allowed but marked noindex · Perplexity-User: Page allowed but marked noindex · Bingbot: Page allowed but marked noindex · Meta-ExternalAgent: Page allowed but marked noindex · Amazonbot: Page allowed but marked noindex · Bytespider: Page allowed but marked noindex · CCBot: Page allowed but marked noindex
  • high
    Page is noindexSearch engines and AI search indexes will drop it. Assistants can still fetch it live.

Every agent

AgentTyperobots.txtFetchWordsResult
Human (Chrome)YouHumanallowed200979 ms111visible
HTTP 200, 111 words
GooglebotGoogle · Classic search index. Also feeds AI Overviews.Search engineallowedAllow: /200982 ms6partial
Page allowed but marked noindex
Google-ExtendedGoogle · Robots-only token controlling Gemini training use. Never fetches on its own.AI trainingallowedAllow: /n/avisible
Robots-only token; allowed by robots.txt
GPTBotOpenAI · Training crawler for OpenAI models.AI trainingallowedAllow: /2001010 ms111partial
Page allowed but marked noindex
OAI-SearchBotOpenAI · Indexes for ChatGPT search results and citations.AI search indexallowedAllow: /200963 ms756partial
Different content than humans receive
ChatGPT-UserOpenAI · Live fetch when a user asks ChatGPT about a page.AI assistant (live fetch)allowedAllow: /200958 ms756partial
Different content than humans receive
ClaudeBotAnthropic · Training crawler for Claude models.AI trainingallowedAllow: /2001005 ms111partial
Page allowed but marked noindex
Claude-SearchBotAnthropic · Indexes for Claude web search.AI search indexallowedAllow: /200980 ms6partial
Page allowed but marked noindex
Claude-UserAnthropic · Live fetch when a user asks Claude about a page.AI assistant (live fetch)allowedAllow: /200959 ms111partial
Page allowed but marked noindex
PerplexityBotPerplexityAI search indexallowedAllow: /200950 ms6partial
Page allowed but marked noindex
Perplexity-UserPerplexityAI assistant (live fetch)allowedAllow: /200952 ms111partial
Page allowed but marked noindex
BingbotMicrosoft · Bing index. Also feeds Copilot and ChatGPT search.Search engineallowed200940 ms6partial
Page allowed but marked noindex
Applebot-ExtendedApple · Robots-only token for Apple Intelligence training.AI trainingallowedn/avisible
Robots-only token; allowed by robots.txt
Meta-ExternalAgentMetaAI trainingallowed200972 ms6partial
Page allowed but marked noindex
AmazonbotAmazon · Alexa and Amazon AI.AI trainingallowed200949 ms6partial
Page allowed but marked noindex
BytespiderByteDance · Known for ignoring robots.txt.AI trainingallowed200978 ms6partial
Page allowed but marked noindex
CCBotCommon Crawl · Open dataset used to train most LLMs.AI trainingallowed200935 ms6partial
Page allowed but marked noindex

Snapshot

Measured 2026-09-25 09:17 UTC from a residential connection. Sites change their rules often, and sites that verify crawler IPs answer differently to different networks. Run a live comparison to see what digicert.com does right now, from Cloudflare's network.

← youtu.be (65)All 100 top 100yuga.com (65) →