AI crawler access:
health publishers
Medical reference sites, where being the cited source carries the highest stakes. 16 domains measured; 7 block at least one of the fifteen tracked AI crawlers.
7of 16 health publishers block at least one AI crawler · median access 100%
Every domain in this category
Sorted by the share of tracked AI crawlers permitted, lowest first. Each report shows all fifteen verdicts and the robots.txt rule behind them.
| Domain | Access | Allowed | Blocked | Not listed |
|---|---|---|---|---|
| verywellhealth.com | 27% | 4 | 11 | 0 |
| drugs.com | 53% | 8 | 7 | 0 |
| healthline.com | 53% | 8 | 7 | 0 |
| medicalnewstoday.com | 53% | 8 | 7 | 0 |
| strava.com | 73% | 11 | 4 | 0 |
| webmd.com | 73% | 11 | 4 | 0 |
| myfitnesspal.com | 93% | 14 | 1 | 0 |
| calm.com | 100% | 15 | 0 | 0 |
| clevelandclinic.org | 100% | 15 | 0 | 0 |
| fitbit.com | 100% | 15 | 0 | 0 |
| goodrx.com | 100% | 15 | 0 | 0 |
| headspace.com | 100% | 15 | 0 | 0 |
| hopkinsmedicine.org | 100% | 15 | 0 | 0 |
| nhs.uk | 100% | 15 | 0 | 0 |
| ouraring.com | 100% | 15 | 0 | 0 |
| teladoc.com | 100% | 0 | 0 | 15 |
Method and caveats are described on the full study page: one robots.txt fetch per domain, resolved against fifteen published AI user-agents, reported as the voluntary policy it is. Check any site yourself with the free crawler access checker.