urlsuma
Listed onlyCrawler for urlsuma.de, a German general-purpose web search engine under construction. The operator documents that it re-reads robots.txt without caching before every single fetch, makes about two requests per visit (robots.txt plus one page), executes no JavaScript, and treats a 4xx response as a permanent block. Current crawler addresses are shared on request only.
Operated by urlsuma.de (Gerhard Stöbe) · Official documentation
User agents
Patterns this directory matches, with real observed strings.
- regexurlsuma/[\d.]+
- regexUrlSuMa\.de crawler
- observedMozilla/5.0 (compatible; urlsuma/2.0; +https://urlsuma.de/bot.html)
- observedMozilla/5.0 (compatible; UrlSuMa.de crawler)
How to verify
No verification recipe published by the operator.
Good Bot Practices scorecard
- Identifies honestly
Stable UA token documented (2 patterns)
- Verifiable
Operator publishes no verification path
- Respects robots.txt
Honors robots.txt (token: urlsuma)
- Behaves
Crawl-rate behavior is operator-declared; not machine-verifiable from this dataset
- Reachable operator
Operator and documentation published
Measured against the Good Bot Practices.