Directory / Search engines

Openindex Spider

Listed only

Openindex operates a web-crawling cluster (Apache Nutch on an Apache Hadoop cluster) for research and development of universal and focused search engines. The crawler respects robots.txt and the Crawl-delay directive and identifies itself in the User-Agent.

Operated by Openindex · Official documentation

User agents

Patterns this directory matches, with real observed strings.

  • regexOpenindexSpider
  • observedMozilla/5.0 (compatible; OpenindexSpider; +https://www.openindex.io/saas/about-our-spider/)

How to verify

No verification recipe published by the operator.

Good Bot Practices scorecard

  • Identifies honestly

    Stable UA token documented (1 pattern)

  • Verifiable

    Operator publishes no verification path

  • Respects robots.txt

    Honors robots.txt (token: OpenindexSpider)

  • Behaves

    Crawl-rate behavior is operator-declared; not machine-verifiable from this dataset

  • Reachable operator

    Operator and documentation published

Measured against the Good Bot Practices.