Tried hard to find difference between pdfs returning no highlighter and ones that do for same search term. Includes pdfs that have been OCRed and ones that were text to begin with. Head scratching to me.
-----Original Message----- From: Erick Erickson [mailto:erickerick...@gmail.com] Sent: Saturday, 10 June 2017 6:22 a.m. To: solr-user <solr-user@lucene.apache.org> Subject: Re: Highlighter not working on some documents Need lots more information. I.e. schema definitions, query you use, handler configuration and the like. Note that highlighted fields must have stored="true" set and likely the _text_ field doesn't. At least in the default schemas stored is set to false for the catch-all field. And you don't want to store that information anyway since it's usually the destination of copyField directives and you'd highlight _those_ fields. Best, Erick On Thu, Jun 8, 2017 at 8:37 PM, Phil Scadden <p.scad...@gns.cri.nz> wrote: > Do a search with: > fl=id,title,datasource&hl=true&hl.method=unified&limit=50&page=1&q=pre > ssure+AND+testing&rows=50&start=0&wt=json > > and I get back a good list of documents. However, some documents are > returning empty fields in the highlighter. Eg, in the highlight array have: > "W:\\Reports\\OCR\\4272.pdf":{"_text_":[]} > > Getting this well up the list of results with good highlighted matchers above > and below this entry. Why would the highlighter be failing? > > Notice: This email and any attachments are confidential and may not be used, > published or redistributed without the prior written consent of the Institute > of Geological and Nuclear Sciences Limited (GNS Science). If received in > error please destroy and immediately notify GNS Science. Do not copy or > disclose the contents. Notice: This email and any attachments are confidential and may not be used, published or redistributed without the prior written consent of the Institute of Geological and Nuclear Sciences Limited (GNS Science). If received in error please destroy and immediately notify GNS Science. Do not copy or disclose the contents.