IndexHalo
MEASURED STUDY · 21 SEPTEMBER 2026

AI crawler access:
developer references

Documentation and Q&A sites — the sources answer engines lean on hardest for technical queries. 34 domains measured; 3 block at least one of the 28 tracked AI crawlers.

3of 34 developer references block at least one AI crawler · median access 100%

Every domain in this category

Sorted by the share of tracked AI crawlers permitted, lowest first. Each report shows all 28 verdicts and the robots.txt rule behind them.

DomainAccessAllowedBlockedNot listed
geeksforgeeks.org89%2530
tutorialspoint.com89%2530
github.com96%2710
angular.io100%2800
codecademy.com100%2800
crates.io100%2800
css-tricks.com100%2800
dev.to100%2800
developer.mozilla.org100%2800
digitalocean.com100%2800
docker.com100%2800
freecodecamp.org100%2800
gitlab.com100%2800
golang.org100%2800
hashnode.com100%2800
kubernetes.io100%2800
leetcode.com100%2800
mongodb.com100%2800
mysql.com100%2800
nodejs.org100%2800
packagist.org100%2800
php.net100%2800
postgresql.org100%2800
pypi.org100%2800
python.org100%2800
reactjs.org100%2800
redis.io100%2800
rubygems.org100%2800
sitepoint.com100%2800
smashingmagazine.com100%2800
sqlite.org100%2800
svelte.dev100%2800
vuejs.org100%2800
w3schools.com100%2800

Method and caveats are described on the full study page: one robots.txt fetch per domain, resolved against 28 published AI user-agents, reported as the voluntary policy it is. Check any site yourself with the free crawler access checker.

Who started blocking which crawler this week, measured from live robots.txt. Unsubscribe in one click.