Openindex Spider
Listed onlyOpenindex operates a web-crawling cluster (Apache Nutch on an Apache Hadoop cluster) for research and development of universal and focused search engines. The crawler respects robots.txt and the Crawl-delay directive and identifies itself in the User-Agent.
Operated by Openindex · Official documentation
User agents
Patterns this directory matches, with real observed strings.
- regexOpenindexSpider
- observedMozilla/5.0 (compatible; OpenindexSpider; +https://www.openindex.io/saas/about-our-spider/)
How to verify
No verification recipe published by the operator.
Good Bot Practices scorecard
- Identifies honestly
Stable UA token documented (1 pattern)
- Verifiable
Operator publishes no verification path
- Respects robots.txt
Honors robots.txt (token: OpenindexSpider)
- Behaves
Crawl-rate behavior is operator-declared; not machine-verifiable from this dataset
- Reachable operator
Operator and documentation published
Measured against the Good Bot Practices.