mirror of
https://github.com/zotero/zotero.git
synced 2026-10-11 03:38:25 +00:00
Every embedding a model produces shares a large common direction that says nothing about the text, so items with little content scored as moderate matches against everything: for the query "sun", an item titled "C" outscored a paper about a coronal mass ejection. Removing that direction spreads the scores out, so no relevance reads as no score. The mean is a constant per model, computed over titles and abstracts across fields and languages, and applies to stored vectors as they're compared, so the index doesn't change. Display ranges are refit to the scores that result. |
||
|---|---|---|
| .. | ||
| components | ||
| content | ||
| resource | ||
| tests | ||
| chrome.manifest | ||
| runtests.sh | ||