IndexHalo
MEASURED STUDY · 21 SEPTEMBER 2026

AI crawler access:
health publishers

Medical reference sites, where being the cited source carries the highest stakes. 16 domains measured; 7 block at least one of the 28 tracked AI crawlers.

7of 16 health publishers block at least one AI crawler · median access 100%

Every domain in this category

Sorted by the share of tracked AI crawlers permitted, lowest first. Each report shows all 28 verdicts and the robots.txt rule behind them.

DomainAccessAllowedBlockedNot listed
verywellhealth.com25%7210
drugs.com61%17110
healthline.com68%1990
medicalnewstoday.com68%1990
strava.com86%2440
webmd.com89%2530
myfitnesspal.com96%2710
calm.com100%2800
clevelandclinic.org100%2800
fitbit.com100%2800
goodrx.com100%2800
headspace.com100%2800
hopkinsmedicine.org100%2800
nhs.uk100%2800
ouraring.com100%2800
teladoc.com100%0028

Method and caveats are described on the full study page: one robots.txt fetch per domain, resolved against 28 published AI user-agents, reported as the voluntary policy it is. Check any site yourself with the free crawler access checker.

Who started blocking which crawler this week, measured from live robots.txt. Unsubscribe in one click.