IndexHalo
MEASURED STUDY · 13 AUGUST 2026

AI crawler access:
health publishers

Medical reference sites, where being the cited source carries the highest stakes. 16 domains measured; 7 block at least one of the fifteen tracked AI crawlers.

7of 16 health publishers block at least one AI crawler · median access 100%

Every domain in this category

Sorted by the share of tracked AI crawlers permitted, lowest first. Each report shows all fifteen verdicts and the robots.txt rule behind them.

DomainAccessAllowedBlockedNot listed
verywellhealth.com27%4110
drugs.com53%870
healthline.com53%870
medicalnewstoday.com53%870
strava.com73%1140
webmd.com73%1140
myfitnesspal.com93%1410
calm.com100%1500
clevelandclinic.org100%1500
fitbit.com100%1500
goodrx.com100%1500
headspace.com100%1500
hopkinsmedicine.org100%1500
nhs.uk100%1500
ouraring.com100%1500
teladoc.com100%0015

Method and caveats are described on the full study page: one robots.txt fetch per domain, resolved against fifteen published AI user-agents, reported as the voluntary policy it is. Check any site yourself with the free crawler access checker.

Who started blocking which crawler this week, measured from live robots.txt. Unsubscribe in one click.