Crawlers
Every search engine, AI model and SEO tool that reads this site, across the website, the API and the image CDN. Updated every hour.
Scoped to cdn.mangabaka.dev · every host
The period being compared against reaches back further than our records, which start 2026-08-28. Changes marked * read larger than they were.
By what they are for
- AI Crawler 1.3M 59.7 GB
- Search engines 53.9k 7.9 GB
- Archiver 5.2k 273 MB
- SEO tooling 3.1k 188 MB
- Page Preview 2.6k 59.4 MB
- Security 980 25.5 MB
- AI search 369 98.2 MB
- AI assistant 139 2.5 MB
- Other 22 437 KB
- Webhooks 4 1.6 MB
AI Crawler are 95% of automated traffic.
Which provider hits it
- Meta 803k 35.2 GB
- Anthropic 217k 9.1 GB
- OpenAI 216k 11.3 GB
- Microsoft 46.5k 7.5 GB
- Amazon 44.1k 4.3 GB
- Internet Archive 5.2k 273 MB
- Ahrefs 2.4k 177 MB
- Google 1.9k 68.0 MB
- Baidu 1.7k 74.5 MB
- DuckDuckGo 1.7k 37.9 MB
Meta is 60% of its requests.
What they take away
- txt 586k 566 MB
- jpeg 371k 19.4 GB
- avif 192k 3.2 GB
- webp 118k 25.8 GB
- png 69.0k 19.3 GB
- json 6.3k 6.0 MB
- empty 450 440 KB
- gif 412 9.2 MB
- html 77 332 KB
Images are 56% of requests and 99% of bytes.
What we answered
- 200 OK 732k 63.8 GB
- 403 refused 585k 565 MB
- 206 partial 17.4k 3.8 GB
- 502 bad gateway 4.1k 4.1 MB
- 404 gone 3.4k 28.6 MB
- 304 unchanged 357 440 KB
- 499 hung up 93 0 B
- 301 moved 34 32.1 KB
44% of bot requests are refused or errored.
Who crawls us
by requestsRequests per day
cdn.mangabaka.devBytes per day
cdn.mangabaka.devCounts come from Cloudflare's edge across both zones, refreshed every hour. Only bots Cloudflare has verified are named, so anything claiming to be Googlebot without the network to prove it is left out. Our own monitoring is excluded.