Directory / Search engines

urlsuma

Listed only

Crawler for urlsuma.de, a German general-purpose web search engine under construction. The operator documents that it re-reads robots.txt without caching before every single fetch, makes about two requests per visit (robots.txt plus one page), executes no JavaScript, and treats a 4xx response as a permanent block. Current crawler addresses are shared on request only.

Operated by urlsuma.de (Gerhard Stöbe) · Official documentation

User agents

Patterns this directory matches, with real observed strings.

  • regexurlsuma/[\d.]+
  • regexUrlSuMa\.de crawler
  • observedMozilla/5.0 (compatible; urlsuma/2.0; +https://urlsuma.de/bot.html)
  • observedMozilla/5.0 (compatible; UrlSuMa.de crawler)

How to verify

No verification recipe published by the operator.

Good Bot Practices scorecard

  • Identifies honestly

    Stable UA token documented (2 patterns)

  • Verifiable

    Operator publishes no verification path

  • Respects robots.txt

    Honors robots.txt (token: urlsuma)

  • Behaves

    Crawl-rate behavior is operator-declared; not machine-verifiable from this dataset

  • Reachable operator

    Operator and documentation published

Measured against the Good Bot Practices.