The Crawl Status window shows how the crawl of your site is going. Open it with Check Crawl Status on the Data Crawler tab. It also opens by itself after you start a crawl.
At the top
- N pages crawled, with Crawling... while pages are still waiting, or Idle when the crawl is done.
- A progress bar and, while the crawl runs, "N remaining in queue, P% complete".
The figures
- In Solr Index: documents of your site's domain now searchable in your Opensolr Index.
- Pages crawled: pages read and kept.
- In Solr buffer: pages read and waiting to be written to the index.
- Remaining in queue: addresses still to read.
- Pages in error: addresses that answered with an error.
- Oversize pages: files larger than the maximum file size of your plan.
- Noindex pages: pages left out on purpose, for example by a noindex instruction.
- Downloaded: how many MB the crawl downloaded.
- Blocked scope: addresses outside the part of the site the crawl may read.
- Redirect outside: addresses that redirect to another site.
Under the figures, Error URLs, Noindex Pages, Blocked Scope and Redirect Outside open the list of addresses behind each count.
Refresh
The window refreshes itself every 5 seconds. Change it with the menu at the top: Auto-refresh: Off, Every 5s, Every 10s, Every 30s or Every 60s. Your browser remembers the choice. The round arrow refreshes at once.
The same figures, and the crawl log, in the Opensolr Control Panel: Crawl Stats and log.