Directory / AI crawlers

ICC-Crawler

Verifiable

ICC-Crawler is a web crawler operated by Japan's NICT that collects web pages across the internet to build datasets for information and language processing research.

Operated by National Institute of Information and Communications Technology (NICT) · Official documentation

User agents

Patterns this directory matches, with real observed strings.

  • regexICC-Crawler
  • observedICC-Crawler/2.0 (Mozilla-compatible; ; http://ucri.nict.go.jp/en/icccrawler.html)

How to verify

Documented static ranges: 202.180.34.186/32, 61.86.246.72/32

IP ranges (2)

Refreshed daily from the operator's feed. Also available as /data/ips/nict-crawler.ips and /data/bots/nict-crawler.json.

202.180.34.186/32
61.86.246.72/32

Good Bot Practices scorecard

  • Identifies honestly

    Stable UA token documented (1 pattern)

  • Verifiable

    Verifiable

  • Respects robots.txt

    Honors robots.txt

  • Behaves

    Crawl-rate behavior is operator-declared; not machine-verifiable from this dataset

  • Reachable operator

    Operator and documentation published

Measured against the Good Bot Practices.