Core Fields of the Site Search Schema

Content, metadata, coordinates, sentiment and search-only fields

Core fields

These fields exist on every document of a site search index. They are what you search, filter, sort and show when you build your own search page. The types are explained in How text is analysed.

Content

FieldTypeWhat it is
idstringThe unique key: the MD5 of the uri (Fields Opensolr adds).
uritext_generalThe address of the page or file, searchable word by word, highlighted.
titletext_generalThe title. It counts most in the default field weights. Highlighted.
descriptiontext_generalThe short description, shown as the result snippet. Highlighted.
texttext_generalThe full body, the largest field. Highlighted.
authortext_generalThe author, searchable; an exact copy is kept in author_s.

How much each field counts is set in Field weights.

Metadata, time and numbers

FieldTypeWhat it is
content_typestringThe file type: text/html, application/pdf and so on. For facets and filters.
content_statusstringThe HTTP status the page answered with when it was read.
signaturestringA hash of the content, used to find duplicates.
categorystringA free category label. For facets and filters.
og_imagestringThe address of the thumbnail. Returned, but not searchable.
creation_datepdateThe publishing date. For date ranges and the freshest results first.
timestampplongThe same moment as Unix time, for fast sorting.
rankpintA ranking value of your own, for your boost functions.
sizepintThe size of the content, 0 by default.

Coordinates and sentiment

FieldTypeWhat it is
coordslocationA latitude,longitude point for distance filters and sorting ({!geofilt}, geodist()).
lat, lonpdoubleThe same coordinates as plain numbers, 0 by default.
sent_pos, sent_neu, sent_negpdoubleHow positive, neutral and negative the text reads, from 0 to 1.
sent_compdoubleThe overall sentiment, from -1 to 1. fq=sent_com:[0.5 TO *] keeps clearly positive content.

Search-only fields (searched, never returned)

FieldTypeWhat it is
spelltextSpellThe dictionary behind "Did you mean".
tags, title_tagsedgy_text_kwMatching the start of a whole phrase, for suggestions while typing.
tags_ws, title_tags_wsedgy_text_wsMatching the start of each word.
embeddingsknn_vectorThe meaning vector of 1,024 numbers behind search by meaning and hybrid search. Not stored: it is searched, never read back.
_version_, _root_plong, stringSolr's own bookkeeping. Leave them alone.

What a page read by the crawl also carries

  • meta_ plus the name of every meta tag of the page, for example meta_og_site_name, along with meta_domain, meta_icon and meta_detected_language.
  • price_f and currency_s: the price found on a product page; product_category_sm: its categories (Prices and product categories).
  • The structured data of the page (JSON-LD, microdata, breadcrumbs), each value in a list field typed by what it holds: _sm, _fm, _im or _dtm.
  • text_t: the structured data as searchable text; video_duration_i and other video fields on pages with a video (Video details); quality_f: the content quality score.

The exact fields of your index are listed in the Control Panel when you set up Facets.

Schema reference pages

Back to Enterprise Site Search