IndexHalo
MEASURED 21 SEPTEMBER 2026

AI crawler access
for pcmag.com

pcmag.com blocks 7 of the 13 crawlers that can cite a source in an AI answer (OAI-SearchBot, ChatGPT-User, PerplexityBot, Perplexity-User, Meta-ExternalFetcher, DuckAssistBot, YouBot). Content behind those rules is unlikely to appear as a cited source on those surfaces.

39%of 28 tracked AI crawlers can reach pcmag.com — 11 allowed, 17 blocked, 0 not listed

That is below the 100% median across the 379 sites in this dataset. Resolved from https://pcmag.com/robots.txt on 21 September 2026.

Every crawler, and the rule that decided it

A crawler named directly in its own group is governed by that group. One not named falls back to the wildcard group. Where neither exists the crawler is reported as not listed — no policy has been expressed, and in practice most crawlers will proceed.

User-agentOperatorPurposeAccessDecided by
OAI-SearchBotOpenAIsearch indexingBlockeddirect rule
GPTBotOpenAImodel trainingBlockeddirect rule
ChatGPT-UserOpenAIon-demand fetchBlockeddirect rule
Claude-SearchBotAnthropicsearch indexingAlloweddirect rule
ClaudeBotAnthropicmodel trainingBlockeddirect rule
Claude-UserAnthropicon-demand fetchAlloweddirect rule
PerplexityBotPerplexitysearch indexingBlockeddirect rule
Perplexity-UserPerplexityon-demand fetchBlockeddirect rule
Google-ExtendedGoogleGemini training controlAllowedwildcard policy
BingbotMicrosoftCopilot search indexAllowedwildcard policy
Applebot-ExtendedAppleApple Intelligence controlBlockeddirect rule
Meta-ExternalAgentMetamodel trainingBlockeddirect rule
AmazonbotAmazonmodel trainingBlockeddirect rule
BytespiderByteDancemodel trainingBlockeddirect rule
CCBotCommon Crawltraining datasetBlockeddirect rule
GoogleOtherGoogleresearch and development fetchAllowedwildcard policy
Google-CloudVertexBotGoogleVertex AI on-demand fetchAllowedwildcard policy
Meta-ExternalFetcherMetaon-demand fetchBlockeddirect rule
MistralAI-UserMistral AIon-demand fetchAlloweddirect rule
cohere-aiCoheremodel trainingBlockeddirect rule
DuckAssistBotDuckDuckGoon-demand fetchBlockeddirect rule
YouBotYou.comsearch indexingBlockeddirect rule
TimpibotTimpitraining datasetBlockeddirect rule
AI2BotAllen Institute for AImodel trainingAlloweddirect rule
DiffbotDiffbotstructured-data extractionBlockeddirect rule
ImagesiftBotImageSifttraining datasetAlloweddirect rule
PetalBotHuaweisearch indexingAlloweddirect rule
OmgilibotWebz.iotraining datasetAllowedwildcard policy

What this does and does not tell you

It tells you what pcmag.com states in its robots.txt, which is the file every well-behaved AI crawler consults before fetching. It is a stated preference and a voluntary protocol — not an access control, and not evidence about what any crawler actually did.

It also says nothing about whether the content is useful to an answer engine once fetched. Access is the precondition; extractable passages, visible dates, named authorship and sourced claims decide whether a reachable page is actually cited. That second layer is what the free evidence-readiness report measures.

Finally, robots.txt changes. This page reflects pcmag.com as it was on 21 September 2026. To see the position right now, run it through the live crawler access checker.

Check your own site

The same 28 user-agents, resolved against your live robots.txt, in about five seconds and without an account: AI crawler access checker. If the result is not what you intended, the robots.txt policy builder composes a corrected block, and how to get cited by ChatGPT explains why blocking GPTBot alone does not remove you from ChatGPT answers.

Browse the full study of 379 measured domains, or the technology media category, to see how this compares.

Embed this result

A live badge for pcmag.com, updated whenever this report refreshes. Paste it into a README or site footer:

EMBED HTML
<a href="https://www.indexhalo.com/ai-crawler-report/pcmag.com/">
  <img src="https://www.indexhalo.com/badge/ai-crawler-access/pcmag.com.svg"
       alt="AI crawler access for pcmag.com: 39%" height="20">
</a>
Who started blocking which crawler this week, measured from live robots.txt. Unsubscribe in one click.