nick-boss-tech opened a new pull request, #4967:
URL: https://github.com/apache/solr/pull/4967

   After the Lucene 10 upgrade, `PathHierarchyTokenizer` emits path components 
as sequential tokens rather than same-position tokens, so `ancestor_path` 
queries (e.g. `Books/NonFic`) stopped matching.
   
   This adds `ZeroPositionIncrementFilter` (plus its factory, registered in 
`TokenFilterFactory` services), which rewrites every token after the first to 
position increment 0 so query builders treat the path components as synonyms 
rather than a phrase — restoring ancestor/descendant path matching. The 
`_default` and sample configset schemas pick up the new filter.
   
   Tests:
   - New `TestZeroPositionIncrementFilterFactory` (4 cases) covering sequential 
tokens sharing the first position, already-zero increments left untouched, and 
single-token input.
   - `PathHierarchyTokenizerFactoryTest` updated for the new token-stream 
behavior; test converted to `SolrTestCase` per reviewer feedback on #4953.
   
   Changelog: `changelog/unreleased/SOLR-18134.yml` (fixed)


-- 
This is an automated message from the Apache Git Service.
To respond to the message, please log on to GitHub and use the
URL above to go to the specific comment.

To unsubscribe, e-mail: [email protected]

For queries about this service, please contact Infrastructure at:
[email protected]


---------------------------------------------------------------------
To unsubscribe, e-mail: [email protected]
For additional commands, e-mail: [email protected]

Reply via email to