Crawl Management
Start the crawl schedule once and Opensolr keeps your index in step with your site: it checks for new and changed pages every few minutes, on its own servers.
Start it
- Pick your content types and click Save configuration.
- Click Start Crawl Schedule. The module registers your sitemap with Opensolr and the first crawl starts.
- The box turns to Crawl Schedule: Active.
The buttons
- Start Crawl Schedule / Stop Crawl Schedule: turns the recurring crawl on or off. Stopping asks for a confirmation; after it, no more automatic crawls run.
- Check for Changes: starts a crawl pass right away. Use it after you publish something you want searchable now.
- Stop Crawl: stops the crawl that is running at this moment. The schedule stays as it is.
- Reindex Everything, Reindex From Scratch, Reset Index: see Reindex and reset.
- Crawl Status: opens the live numbers. See Crawl Status.
When an editor saves a page
When a published page of a crawled type is saved, the module takes it out of the index and out of the crawl list at once, so the next crawl pass reads it fresh. An unpublished or deleted page leaves the index right away. With real-time sync on, the saved page is also sent straight away, without waiting for the crawl.
Above the buttons, Crawl Queue counts the addresses in your sitemap: pages, files and the number of sitemaps they are split into.
Data Crawler
- Data Crawler
- Content Types for Crawling
- Crawler Settings
- Crawl Management
- Reindex and Reset
- Crawl Status