Crawler documentation
The DiscoverBot.
Finding new paths. WebAtlasDiscoverBot explores the web to feed our global indexing queue.
Crawler documentation
Finding new paths. WebAtlasDiscoverBot explores the web to feed our global indexing queue.
WebAtlasDiscoverBot identifies itself with the following User-Agent.
WebAtlasDiscoverBot/2.1 (+https://webatlasindex.pl/discover-bot.php)
A lightweight crawler focused purely on link discovery. It respects robots.txt rules under its own name, downloads at most 512 KB per connection, and honors Crawl-delay to stay low-impact on your server.
Every discovered link passes through three filters before it's queued:
/admin, /wp-login, /api).gclid, utm_source, clickid).To block WebAtlasDiscoverBot entirely, add this to your robots.txt:
User-agent: WebAtlasDiscoverBot
Disallow: /