docs: explain the Constella family and recommended retrieval setup
Browse files
README.md
CHANGED
|
@@ -25,6 +25,9 @@ It produces normalized 1024-dimensional vectors that search documents encoded by
|
|
| 25 |
The same document index also works with the stronger
|
| 26 |
[`constella-nano`](https://huggingface.co/DylanCouzon/constella-nano) query encoder.
|
| 27 |
|
|
|
|
|
|
|
|
|
|
| 28 |
| Property | Value |
|
| 29 |
|---|---|
|
| 30 |
| Role | Query encoder |
|
|
@@ -34,6 +37,17 @@ The same document index also works with the stronger
|
|
| 34 |
| Maximum input length | 512 tokens |
|
| 35 |
| Query prefix | None |
|
| 36 |
| Document encoder | `DylanCouzon/stella-en-400M-v5-doc-onnx` |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 37 |
|
| 38 |
## Installation
|
| 39 |
|
|
@@ -134,11 +148,12 @@ therefore be interpreted separately from the other four.
|
|
| 134 |
|
| 135 |
Note: Stella discloses training or evaluation contact with ArguAna and FiQA.
|
| 136 |
|
| 137 |
-
|
| 138 |
-
|
| 139 |
-
|
| 140 |
-
|
| 141 |
-
BM25 implementation.
|
|
|
|
| 142 |
|
| 143 |
Zero did not establish an improvement over BM25 in the six-dataset statistical test. Its measured
|
| 144 |
difference was +0.0165 nDCG@10, but the adjusted test threshold was not met. This is not an
|
|
|
|
| 25 |
The same document index also works with the stronger
|
| 26 |
[`constella-nano`](https://huggingface.co/DylanCouzon/constella-nano) query encoder.
|
| 27 |
|
| 28 |
+
> Research preview: Native FastEmbed support currently requires the Constella preview branch
|
| 29 |
+
> shown below. The published evaluation is limited to the results described in this card.
|
| 30 |
+
|
| 31 |
| Property | Value |
|
| 32 |
|---|---|
|
| 33 |
| Role | Query encoder |
|
|
|
|
| 37 |
| Maximum input length | 512 tokens |
|
| 38 |
| Query prefix | None |
|
| 39 |
| Document encoder | `DylanCouzon/stella-en-400M-v5-doc-onnx` |
|
| 40 |
+
| Recommended retrieval | Hybrid with BM25 and DBSF at prefetch 100 |
|
| 41 |
+
|
| 42 |
+
## The Constella family
|
| 43 |
+
|
| 44 |
+
The name Constella combines "constellation" and "Stella." The document embeddings are the fixed
|
| 45 |
+
stars, and the query encoder navigates their shared vector space.
|
| 46 |
+
|
| 47 |
+
Zero and Nano are swappable at query time. Both can search the same document index, so you can
|
| 48 |
+
choose between them without re-encoding documents or rebuilding the collection. They do not
|
| 49 |
+
produce identical rankings: Zero is the faster option, while Nano has higher retrieval scores on
|
| 50 |
+
the six reported datasets. The "zero" name refers to its transformer-free query path.
|
| 51 |
|
| 52 |
## Installation
|
| 53 |
|
|
|
|
| 148 |
|
| 149 |
Note: Stella discloses training or evaluation contact with ArguAna and FiQA.
|
| 150 |
|
| 151 |
+
The recommended deployment setup for Zero is hybrid retrieval. Retrieve with both Zero and BM25,
|
| 152 |
+
then combine their results with Qdrant's distribution-based score fusion (DBSF), prefetching 100
|
| 153 |
+
candidates from each side. This setup scored 0.4887 mean nDCG@10 across all six datasets and
|
| 154 |
+
0.4912 across the four datasets without disclosed Stella contact. The evaluated lexical side used
|
| 155 |
+
`bm25s` with Lucene defaults, so results may differ with another BM25 implementation. Dense-only
|
| 156 |
+
retrieval remains supported when a lexical index is unavailable or unnecessary.
|
| 157 |
|
| 158 |
Zero did not establish an improvement over BM25 in the six-dataset statistical test. Its measured
|
| 159 |
difference was +0.0165 nDCG@10, but the adjusted test threshold was not met. This is not an
|