Crawler documentation
The RssSitemapBot.
Reads RSS, Atom, and sitemap files to feed our global indexing queue.
Crawler documentation
Reads RSS, Atom, and sitemap files to feed our global indexing queue.
WebAtlasRssSitemapBot identifies itself with the following User-Agent.
WebAtlasRssSitemapBot/1.1 (+https://webatlasindex.pl/rsssitemap-bot.php)
A lightweight crawler whose job is to read submitted RSS/Atom feeds and sitemap files, then extract the URLs they contain.
Extracted URLs are automatically filtered through three layers:
/admin, /wp-login, /api).gclid, utm_source, clickid).To block WebAtlasRssSitemapBot entirely, add this to your robots.txt:
User-agent: WebAtlasRssSitemapBot
Disallow: /