<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>Verified Bots Directory — new bots</title>
    <link>https://verifiedbots.dev</link>
    <atom:link href="https://verifiedbots.dev/rss.xml" rel="self" type="application/rss+xml" />
    <description>Bots newly added to the directory, with identity, category, and verification tier.</description>
    <language>en</language>
    <lastBuildDate>Sat, 05 Sep 2026 04:31:23 GMT</lastBuildDate>
    <item>
      <title>Fedicabot — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/fedicabot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/fedicabot</guid>
      <pubDate>Thu, 03 Sep 2026 13:20:13 GMT</pubDate>
      <description>The fetcher used by Fedica's social-media publishing platform. The operator documents that it does not crawl whole sites: it reads only the page or RSS feed a Fedica user has specified, downloading feed content so the user can schedule posts from it, and reading Open Graph metadata so a shared link renders as a card. Fedica publishes the robots.txt token but states no robots.txt policy and no address ranges. Operated by Fedica.</description>
    </item>
    <item>
      <title>Swiftbot — Search engines, listed only</title>
      <link>https://verifiedbots.dev/bots/swiftbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/swiftbot</guid>
      <pubDate>Thu, 03 Sep 2026 13:20:13 GMT</pubDate>
      <description>The web crawler behind Swiftype's site-search product, now operated by Elasticsearch B.V. Unlike a general search-engine crawler it only visits sites that Swiftype customers have asked it to crawl, in order to build a search index for those sites. The operator documents that Swiftbot obeys every restriction it finds in robots.txt, including a full-site disallow, and that it looks for the token Swiftbot. Operated by Elasticsearch B.V. (Swiftype).</description>
    </item>
    <item>
      <title>OAI-AdsBot — SEO tools, fully verifiable</title>
      <link>https://verifiedbots.dev/bots/oai-adsbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/oai-adsbot</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>OpenAI's fetcher for advertising on ChatGPT. When an advertiser submits an ad, OAI-AdsBot visits the submitted landing page to check that it complies with OpenAI's ad policies and to judge when the ad is relevant to show. OpenAI states it only visits pages submitted as ads and that what it collects is not used to train generative AI foundation models. It has its own published IP range list, separate from GPTBot, OAI-SearchBot and ChatGPT-User. Operated by OpenAI.</description>
    </item>
    <item>
      <title>W3C Link Checker — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/w3c-checklink</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/w3c-checklink</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>W3C's link checker, which follows the links on a user-submitted page to report broken and redirected references. W3C's validation services page marks it as the one validator that, being a crawling service, honors robots.txt directives, and gives the same published source addresses as the other W3C validators. Operated by World Wide Web Consortium (W3C).</description>
    </item>
    <item>
      <title>W3C CSS Validation Service — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/w3c-css-validator</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/w3c-css-validator</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>W3C's CSS validator, which fetches a user-submitted page and its stylesheets to check them against the CSS specifications. It runs on W3C's Jigsaw server and identifies itself with a Jigsaw prefix followed by its own W3C_CSS_Validator_JFouffa token, from the same published validator addresses. Operated by World Wide Web Consortium (W3C).</description>
    </item>
    <item>
      <title>W3C Feed Validation Service — Feed fetchers, verifiable</title>
      <link>https://verifiedbots.dev/bots/w3c-feed-validator</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/w3c-feed-validator</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>W3C's feed validator, which fetches a user-submitted RSS or Atom feed to check it against the feed specifications. It is a one-off conformance check rather than a subscription fetcher, and W3C publishes its user agent and the shared validator source addresses. Operated by World Wide Web Consortium (W3C).</description>
    </item>
    <item>
      <title>W3C Internationalization Checker — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/w3c-i18n-checker</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/w3c-i18n-checker</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>The fetcher behind W3C's Internationalization Checker, which retrieves a user-submitted page and reports on its language, encoding and other i18n markup. W3C's validation services page lists its user agent alongside the single IPv4 address and IPv6 prefix all W3C validators fetch from. Operated by World Wide Web Consortium (W3C).</description>
    </item>
    <item>
      <title>Validator.nu (W3C Nu HTML Checker) — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/w3c-validator-nu</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/w3c-validator-nu</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>The fetcher behind the Nu HTML Checker, the W3C validation service that checks a user-submitted URL against current HTML conformance rules. W3C's validation services page lists its user agent and states that all W3C validator traffic comes from one published IPv4 address and one IPv6 prefix. Operated by World Wide Web Consortium (W3C).</description>
    </item>
    <item>
      <title>YandexAccessibilityBot — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/yandex-accessibility</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yandex-accessibility</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>Yandex's accessibility checker, which downloads pages to assess how accessible they are to users. Yandex documents that it sends up to three requests per second, ignores the crawl-rate setting in Yandex Webmaster, and does not follow the general robots.txt rules written for arbitrary robots. Operated by Yandex.</description>
    </item>
    <item>
      <title>YandexAdditionalBot — AI crawlers, verifiable</title>
      <link>https://verifiedbots.dev/bots/yandex-additional</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yandex-additional</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>Yandex's robot for processing robots.txt so that site owners can keep already indexed page content out of the AI responses shown in Yandex Search. Yandex's robot table states it applies to pages that have already been indexed by the primary crawler and that it makes no indexing requests of its own. Operated by Yandex.</description>
    </item>
    <item>
      <title>YandexComBot — Search engines, verifiable</title>
      <link>https://verifiedbots.dev/bots/yandex-com</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yandex-com</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>Yandex's indexing robot for content in languages other than Russian, feeding the international yandex.com search index. Yandex's robot table documents that it can index content when there is no explicit robot-specific restriction, and lists it among the robots that do not follow the general robots.txt rules written for arbitrary robots. Operated by Yandex.</description>
    </item>
    <item>
      <title>YandexMobileBot — Search engines, verifiable</title>
      <link>https://verifiedbots.dev/bots/yandex-mobile</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yandex-mobile</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>Yandex's mobile-layout robot, which fetches pages to determine whether their layout is suitable for mobile devices. It advertises an iPhone Safari prefix followed by its own YandexMobileBot token, and Yandex lists it among the robots that do not follow the general robots.txt rules written for arbitrary robots. Operated by Yandex.</description>
    </item>
    <item>
      <title>YandexPagechecker — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/yandex-pagechecker</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yandex-pagechecker</guid>
      <pubDate>Tue, 01 Sep 2026 13:17:25 GMT</pubDate>
      <description>The fetcher behind Yandex's structured data validator, which retrieves a page so its markup can be validated. Yandex's robot table records it as taking the general robots.txt rules into account. Operated by Yandex.</description>
    </item>
    <item>
      <title>AlertSite — Monitoring, verifiable</title>
      <link>https://verifiedbots.dev/bots/alertsite</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/alertsite</guid>
      <pubDate>Mon, 31 Aug 2026 13:20:25 GMT</pubDate>
      <description>SmartBear's AlertSite runs customer-configured availability and transaction monitors against the sites and APIs their owners point it at, from a published set of monitoring locations. SmartBear tells site owners to filter the traffic by looking for &quot;AlertSite&quot; in the user agent string or by the monitoring location addresses. Some mobile-network locations use dynamic carrier addresses that are not in the published list. Operated by SmartBear Software.</description>
    </item>
    <item>
      <title>FeedWind Crawler — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/feedwind-crawler</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/feedwind-crawler</guid>
      <pubDate>Mon, 31 Aug 2026 13:20:25 GMT</pubDate>
      <description>The crawler behind FeedWind, an embeddable RSS/Atom widget. It fetches the feed sources that a widget owner has configured, every five minutes to five hours depending on their plan, so the rendered widget stays current. The operator's support page prints the crawler's full user agent and asks feed publishers to unblock it if a firewall is turning it away. Operated by Mikle KK (FeedWind).</description>
    </item>
    <item>
      <title>Ghost Inspector — Monitoring, verifiable</title>
      <link>https://verifiedbots.dev/bots/ghost-inspector</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/ghost-inspector</guid>
      <pubDate>Mon, 31 Aug 2026 13:20:25 GMT</pubDate>
      <description>Ghost Inspector runs customer-authored browser tests against the sites their owners point it at, from a fixed pool of cloud test runners in 18 regions. Its runners append the literal string &quot;Ghost Inspector&quot; to the browser user agent they run, and the operator publishes the full list of test-runner addresses so site owners can filter the traffic out of their analytics. Operated by Ghost Inspector.</description>
    </item>
    <item>
      <title>RepoLookoutBot — Security scanners, listed only</title>
      <link>https://verifiedbots.dev/bots/repo-lookout-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/repo-lookout-bot</guid>
      <pubDate>Mon, 31 Aug 2026 13:20:25 GMT</pubDate>
      <description>Repo Lookout is a non-commercial internet-wide security scanner that looks for source-code repositories accidentally exposed to the public web and reports them to the domain's technical contact. The operator documents the scanner's user-agent prefix and offers two opt-outs: an email request, or denying every request whose user agent starts with RepoLookoutBot. Operated by Crissy Field GmbH (Repo Lookout).</description>
    </item>
    <item>
      <title>WepchSearchEngine — Search engines, listed only</title>
      <link>https://verifiedbots.dev/bots/wepch-searchengine</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/wepch-searchengine</guid>
      <pubDate>Mon, 31 Aug 2026 13:20:25 GMT</pubDate>
      <description>The crawler for Wepch, an independent privacy-focused search engine still in development. Its operator page states the project's purpose, gives a contact address for crawl complaints, and documents the robots.txt token site owners can use to crawl-delay or disallow it. Operated by Wepch.</description>
    </item>
    <item>
      <title>atlassian-bot — Search engines, verifiable</title>
      <link>https://verifiedbots.dev/bots/atlassian-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/atlassian-bot</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Crawler behind the Teamwork Graph &quot;custom website&quot; connector, which indexes a site an Atlassian customer has connected so its content is searchable in Rovo. Atlassian documents the robots.txt token and publishes the connector's egress ranges under the rovo-crawler product in its IP-ranges feed. Operated by Atlassian.</description>
    </item>
    <item>
      <title>EasyCron — Webhooks, fully verifiable</title>
      <link>https://verifiedbots.dev/bots/easycron</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/easycron</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Hosted cron service that calls URLs its users schedule, on the interval or cron expression they configure. Requests go only to endpoints the user entered, and EasyCron publishes the IPv4 and IPv6 addresses its execution bots call from as a JSON list. Operated by EasyCron.</description>
    </item>
    <item>
      <title>EdgeWatch — Security scanners, verifiable</title>
      <link>https://verifiedbots.dev/bots/edgewatch</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/edgewatch</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Internet-wide scanner for EdgeWatch's attack-surface and threat-intelligence platform. It makes connection attempts to public IPv4 and IPv6 addresses daily and completes protocol handshakes to fingerprint the services it finds; EdgeWatch documents the subnet its scanners use and an e-mail opt-out. Operated by EdgeWatch.</description>
    </item>
    <item>
      <title>FedReporterDataBot — Monitoring, listed only</title>
      <link>https://verifiedbots.dev/bots/fedreporter-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/fedreporter-bot</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Scheduled data-retrieval bot that pulls publicly available banking-institution records from the FFIEC National Information Center for Fed Reporter's internal research dashboards. Its documentation states it targets only public data, uses no CAPTCHA bypass, and runs a separate testing agent off-peak. Operated by Fed Reporter Inc..</description>
    </item>
    <item>
      <title>Productsup Website Crawler — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/productsup-crawler</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/productsup-crawler</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Crawler behind Productsup's website data-import source, which reads pages of a site a Productsup customer has configured and turns them into a product data feed. Productsup documents the default user agent, and notes that customers may append a hash to it for their own filtering. Operated by Productsup.</description>
    </item>
    <item>
      <title>rakutenusabot-image — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/rakuten-usa-image-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/rakuten-usa-image-bot</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Rakuten's image-extraction bot, which retrieves product images from merchant sites that have partnered with Rakuten. Its bot page names the identifying token that appears in the user-agent string and gives an abuse contact. Operated by Rakuten.</description>
    </item>
    <item>
      <title>Scrunchbot — Monitoring, listed only</title>
      <link>https://verifiedbots.dev/bots/scrunchbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/scrunchbot</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Fetcher for Scrunch AI's brand-visibility product. Scrunch documents that it accesses customer websites to diagnose issues with their content and configuration, and that it does not proactively crawl or spider the general internet — requests are made only in response to customer configuration. Operated by Scrunch AI.</description>
    </item>
    <item>
      <title>SirdataBot — SEO tools, listed only</title>
      <link>https://verifiedbots.dev/bots/sirdata-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/sirdata-bot</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Crawler for Sirdata's contextual advertising API. It fetches and categorises page content when Sirdata's on-page script cannot read the document from the parent frame. Sirdata publishes the addresses it crawls from at proxies-list.sirdata.fr, but that pool rotates continuously — it recommends re-fetching every ten minutes — so no snapshot of it stays accurate. Operated by Sirdata.</description>
    </item>
    <item>
      <title>XY-Archive-Compliance — Archivers, verifiable</title>
      <link>https://verifiedbots.dev/bots/xy-archive-compliance</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/xy-archive-compliance</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Website archiver for XY Archive, XYPN's compliance-archiving product for financial advisory firms. It crawls a subscribing firm's site to decide which pages to capture and then screenshots each page for the archive; XYPN documents the single address it archives from. Operated by XY Planning Network.</description>
    </item>
    <item>
      <title>Yahoo Ad Monitoring — SEO tools, listed only</title>
      <link>https://verifiedbots.dev/bots/yahoo-ad-monitoring</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yahoo-ad-monitoring</guid>
      <pubDate>Sat, 29 Aug 2026 13:18:16 GMT</pubDate>
      <description>Page-fetch client that retrieves the landing pages of URLs listed with Yahoo advertising services. Yahoo documents that it fetches each advertiser-supplied landing page to check policy compliance and to improve the accuracy of the ad listing, and that it uses separate desktop and mobile user agents. Operated by Yahoo.</description>
    </item>
    <item>
      <title>Deskyobot — Social previews, listed only</title>
      <link>https://verifiedbots.dev/bots/deskyobot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/deskyobot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Link-preview fetcher for Deskyo, a service that lets its members save web pages. The operator's bot page documents that Deskyo must display information tied to a saved page — its title, description and favicon — which is what this fetcher retrieves, and gives a legal contact for questions about links saved by a member. The page states no robots.txt policy and publishes no addresses. Operated by Deskyo.</description>
    </item>
    <item>
      <title>Feeder Crawler — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/feeder-co</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/feeder-co</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Crawler for Feeder, a hosted RSS reading service. The operator documents that it only fetches feeds users have actively subscribed to, that it acts as an agent on the user's behalf and therefore does not check robots.txt, and that rate limits can be arranged by contacting support. Its user agent is a browser-shaped string in which Feeder deliberately inserts the feeder.co identifier as the honest self-identification. Operated by Feeder.</description>
    </item>
    <item>
      <title>FleebsBot — Search engines, listed only</title>
      <link>https://verifiedbots.dev/bots/fleebsbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/fleebsbot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Crawler for fleebs.com, a German real-time search engine. The operator's bot page documents two identities: one that searches for new information and one that analyses the pages it finds. It documents robots.txt blocking with the FleebsBot token and notes that pages indexed before a block may take time to drop out of the index. Operated by iontic GmbH (fleebs.com).</description>
    </item>
    <item>
      <title>Innguma Fetcher — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/innguma-fetcher</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/innguma-fetcher</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Innguma's feed fetcher. It retrieves RSS and Atom feeds that users have added to Innguma or to another application built on the Innguma cloud, and refreshes them roughly once an hour. The operator states that because the requests follow explicit action by human users rather than automated crawling, the fetcher does not follow robots.txt, and documents serving an error status to the Innguma/1.0 user agent as the opt-out. Operated by Innguma.</description>
    </item>
    <item>
      <title>LinkCheckerBot — SEO tools, fully verifiable</title>
      <link>https://verifiedbots.dev/bots/linkcheckerbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/linkcheckerbot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>LinkChecker.pro's backlink-monitoring crawler. It revisits the backlinks that the service's customers track, checking whether each link is still present and how it is marked. The operator documents that the crawler strictly respects robots.txt and Crawl-delay, and publishes the addresses it crawls from. Operated by Local Profy LLC (LinkChecker.pro).</description>
    </item>
    <item>
      <title>McontextualBot — SEO tools, listed only</title>
      <link>https://verifiedbots.dev/bots/mcontextual-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/mcontextual-bot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>MContextual's crawler for contextual advertising. The operator documents that it reads page content so advertisers can build cookieless contextual audiences from a site's subject matter, that it can be blocked with standard robots.txt directives, and that its requests originate from AWS and GCP addresses rather than a published range. Operated by MContextual.</description>
    </item>
    <item>
      <title>MediaMonitoringBot — Monitoring, verifiable</title>
      <link>https://verifiedbots.dev/bots/mediamonitoringbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/mediamonitoringbot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Crawler behind MediaMonitoringBot, a Ukrainian media-monitoring service that indexes news and media publisher pages and delivers mention alerts to PR and communications subscribers. The operator documents that it strictly respects robots.txt allow and disallow rules, publishes its main crawler addresses, and documents a reverse-DNS suffix for verification. Operated by MediaMonitoringBot.</description>
    </item>
    <item>
      <title>Mozilla-Tabstack — AI assistants, listed only</title>
      <link>https://verifiedbots.dev/bots/mozilla-tabstack</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/mozilla-tabstack</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Fetcher for Tabstack, Mozilla's developer-facing platform for programmatic, AI-driven interaction with web content. Every request carries a dedicated user agent, and the operator documents that Tabstack respects robots.txt rules addressed to it, stops immediately on a disallowed path, fails fast rather than retrying, and caches robots.txt results to reduce follow-up requests. Operated by Mozilla (Tabstack).</description>
    </item>
    <item>
      <title>NestDaddyBot — Search engines, listed only</title>
      <link>https://verifiedbots.dev/bots/nestdaddybot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/nestdaddybot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Web crawler behind the NestDaddy search engine. The operator's webmaster page documents that it extracts page title, meta description, headings, language and link structure to build the index, filters adult and abusive content out of results, and fully respects robots.txt including the Crawl-delay directive. It states that the crawler runs from a distributed network whose current ranges are available only on request. Operated by NestDaddy Technologies.</description>
    </item>
    <item>
      <title>ObservePoint — Monitoring, verifiable</title>
      <link>https://verifiedbots.dev/bots/observepoint</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/observepoint</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>ObservePoint's web-governance platform runs customer-configured audits and user journeys against their own sites to validate analytics and marketing tags. Requests come from a published set of static addresses that the operator states are used by ObservePoint alone, and most of its browser-based user agents carry the literal name &quot;ObservePoint&quot;. Operated by ObservePoint.</description>
    </item>
    <item>
      <title>ScourRSSBot — Feed fetchers, listed only</title>
      <link>https://verifiedbots.dev/bots/scour-rss-bot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/scour-rss-bot</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Feed fetcher for Scour, a personalized content feed service in which users subscribe to RSS, Atom and JSON feeds and topics of interest. It polls each subscribed feed once every 900 seconds regardless of subscriber count, uses conditional requests, rate-limits itself to two requests per second per registrable domain, and backs off on error responses. Operated by Scour (Evan Schwartz).</description>
    </item>
    <item>
      <title>urlsuma — Search engines, listed only</title>
      <link>https://verifiedbots.dev/bots/urlsuma</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/urlsuma</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Crawler for urlsuma.de, a German general-purpose web search engine under construction. The operator documents that it re-reads robots.txt without caching before every single fetch, makes about two requests per visit (robots.txt plus one page), executes no JavaScript, and treats a 4xx response as a permanent block. Current crawler addresses are shared on request only. Operated by urlsuma.de (Gerhard Stöbe).</description>
    </item>
    <item>
      <title>Yahoo Mail Proxy — Social previews, verifiable</title>
      <link>https://verifiedbots.dev/bots/yahoo-mail-proxy</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/yahoo-mail-proxy</guid>
      <pubDate>Thu, 27 Aug 2026 13:23:13 GMT</pubDate>
      <description>Yahoo Mail's content-fetch proxy. It retrieves images and other assets embedded in the URLs of messages sent to Yahoo Mail users, so that the content is served through Yahoo rather than fetched directly by the recipient's browser. Yahoo documents the user agent and publishes the address ranges the proxy fetches from. Operated by Yahoo.</description>
    </item>
    <item>
      <title>Artemis Web Reader — Feed fetchers, fully verifiable</title>
      <link>https://verifiedbots.dev/bots/artemis-web-reader</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/artemis-web-reader</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>The feed poller behind Artemis, a hosted web reader that follows websites and blogs on behalf of its subscribers. It updates once per day and uses HEAD and conditional GET requests with If-Modified-Since and If-None-Match to reduce bandwidth. Artemis publishes the addresses it polls feeds from as a plain-text list. Operated by Artemis (jamesg.blog).</description>
    </item>
    <item>
      <title>CLASSLA-web — Archivers, listed only</title>
      <link>https://verifiedbots.dev/bots/classla-web</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/classla-web</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>The CLASSLA-web crawler collects web corpora for South Slavic and other languages under the CLARIN Knowledge Centre for South Slavic languages, continuing the CEF-funded MaCoCu project's crawling infrastructure. It runs the SpiderLing crawler developed at Masaryk University. The operator documents that it adheres to the robots exclusion standard and reads robots.txt on the first access of each crawl run. Operated by CLARIN.SI (CLASSLA Knowledge Centre).</description>
    </item>
    <item>
      <title>Cotoyogi — Archivers, verifiable</title>
      <link>https://verifiedbots.dev/bots/cotoyogi</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/cotoyogi</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>A research web crawler run by the Data Lake Research and Development Center of the Joint Support-Center for Data Science Research (ROIS-DS), Research Organization of Information and Systems, Japan. It collects Japanese-language web data for research use. The operator page documents RFC 9309 robots.txt handling with Crawl-delay support, robots meta tag support, and the address range the crawler requests from. Operated by ROIS-DS Data Lake Research and Development Center.</description>
    </item>
    <item>
      <title>FindFilesBot — Search engines, verifiable</title>
      <link>https://verifiedbots.dev/bots/findfilesbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/findfilesbot</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>The crawler behind the FindFiles.net file search engine. It checks that publicly accessible files users search for are still available, using multiplexed HTTP/2 HEAD requests with a minimum ten-second interval per server, and may download images, videos and executables for classification and safety checks. Purpose-specific user agents are published for its link checking, favicon fetching, virus scanning and image classification agents, which all share one crawler host and address. Operated by FindFiles.net.</description>
    </item>
    <item>
      <title>IbouBot — Search engines, fully verifiable</title>
      <link>https://verifiedbots.dev/bots/iboubot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/iboubot</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>The crawler behind Ibou, a French conversational search engine. It discovers publicly accessible pages and indexes them so Ibou can cite and link back to the sites it answers from. The operator documents a 5-second politeness delay per host, publishes its address ranges as JSON, and documents forward-confirmed reverse DNS as the reliable way to tell a genuine request from a spoofed user agent. Operated by Babbar (Ibou).</description>
    </item>
    <item>
      <title>LinksIndexerBot — SEO tools, verifiable</title>
      <link>https://verifiedbots.dev/bots/linksindexerbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/linksindexerbot</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>The crawler behind the Links Indexer URL indexing service. It parses third-party sites to verify their URLs and status for sitemap campaigns, querying pages for metadata and favicons and taking a homepage screenshot. The operator states it never harvests e-mail addresses or content unrelated to sitemap campaigns, runs at most four parallel requests, and publishes the single address it crawls from. Operated by Links Indexer (Kalpraj Solutions (OPC) Private Limited).</description>
    </item>
    <item>
      <title>LyonlBot — Search engines, listed only</title>
      <link>https://verifiedbots.dev/bots/lyonlbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/lyonlbot</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>The crawler used by Lyonl Search to discover and refresh public web documents for its index, with separate identities for web, image and news crawling so site owners can control each class independently. The operator documents robots.txt and crawl-delay support, conditional fetches with ETag and Last-Modified, and that the crawler does not submit forms or bypass authentication. It states that crawler address ranges are not currently published, so no verification recipe is recorded. Operated by Lyonl Search.</description>
    </item>
    <item>
      <title>MagnetmeBot — Search engines, fully verifiable</title>
      <link>https://verifiedbots.dev/bots/magnetmebot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/magnetmebot</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>Magnet.me's web crawler. It discovers and refreshes career-related pages — vacancies, events and employer information — for the Magnet.me careers platform. The operator documents that it may render JavaScript, follows the canonical version of a page, does not submit forms and adjusts its crawl speed automatically, and it publishes the addresses it crawls from as both a JSON and a plain-text list. Operated by Magnet.me.</description>
    </item>
    <item>
      <title>MotoMinerBot — Search engines, verifiable</title>
      <link>https://verifiedbots.dev/bots/motominerbot</link>
      <guid isPermaLink="true">https://verifiedbots.dev/bots/motominerbot</guid>
      <pubDate>Tue, 25 Aug 2026 13:21:31 GMT</pubDate>
      <description>MotoMiner's vehicle-listing crawler. It primarily crawls automotive dealership websites and adds their vehicle detail pages to the MotoMiner index. The operator documents robots.txt and crawl-delay support, honours bot-specific noindex meta directives, throttles outbound requests against defined traffic thresholds, and publishes the address the crawler currently requests from. Operated by MotoMiner.</description>
    </item>
  </channel>
</rss>
