AI crawler access:
health publishers
Medical reference sites, where being the cited source carries the highest stakes. 16 domains measured; 7 block at least one of the 28 tracked AI crawlers.
7of 16 health publishers block at least one AI crawler · median access 100%
Every domain in this category
Sorted by the share of tracked AI crawlers permitted, lowest first. Each report shows all 28 verdicts and the robots.txt rule behind them.
| Domain | Access | Allowed | Blocked | Not listed |
|---|---|---|---|---|
| verywellhealth.com | 25% | 7 | 21 | 0 |
| drugs.com | 61% | 17 | 11 | 0 |
| healthline.com | 68% | 19 | 9 | 0 |
| medicalnewstoday.com | 68% | 19 | 9 | 0 |
| strava.com | 86% | 24 | 4 | 0 |
| webmd.com | 89% | 25 | 3 | 0 |
| myfitnesspal.com | 96% | 27 | 1 | 0 |
| calm.com | 100% | 28 | 0 | 0 |
| clevelandclinic.org | 100% | 28 | 0 | 0 |
| fitbit.com | 100% | 28 | 0 | 0 |
| goodrx.com | 100% | 28 | 0 | 0 |
| headspace.com | 100% | 28 | 0 | 0 |
| hopkinsmedicine.org | 100% | 28 | 0 | 0 |
| nhs.uk | 100% | 28 | 0 | 0 |
| ouraring.com | 100% | 28 | 0 | 0 |
| teladoc.com | 100% | 0 | 0 | 28 |
Method and caveats are described on the full study page: one robots.txt fetch per domain, resolved against 28 published AI user-agents, reported as the voluntary policy it is. Check any site yourself with the free crawler access checker.