ipipan
/

silver-retriever-base-v1

@@ -26,14 +26,51 @@ Silver Retriever model encodes the Polish sentences or paragraphs into a 768-dim
 It was initialized from the [HerBERT-base](https://huggingface.co/allegro/herbert-base-cased) model and fine-tuned on the [PolQA](https://huggingface.co/ipipan/polqa) and [MAUPQA](https://huggingface.co/ipipan/maupqa) datasets for 15,000 steps with a batch size of 1,024.
-## Preparing inputs
 The model was trained on question-passage pairs and works best when the input is the same format as that used during training:
-- We added the phrase `Pytanie:' to the beginning of the question.
 - The training passages consisted of `title` and `text` concatenated with the special token `</s>`. Even if your passages don't have a `title`, it is still beneficial to prefix a passage with the `</s>` token.
 - Although we used the dot product during training, the model usually works better with the cosine distance.
-## Usage (Sentence-Transformers)
 Using this model becomes easy when you have [sentence-transformers](https://www.SBERT.net) installed:
@@ -55,7 +92,7 @@ embeddings = model.encode(sentences)
 print(embeddings)
 ```
-## Usage (HuggingFace Transformers)
 Without [sentence-transformers](https://www.SBERT.net), you can use the model like this: First, you pass your input through the transformer model, then you have to apply the right pooling-operation on-top of the contextualized word embeddings.
 ```python

 It was initialized from the [HerBERT-base](https://huggingface.co/allegro/herbert-base-cased) model and fine-tuned on the [PolQA](https://huggingface.co/ipipan/polqa) and [MAUPQA](https://huggingface.co/ipipan/maupqa) datasets for 15,000 steps with a batch size of 1,024.
+## Evaluation
+### Accuracy@10
+| **Model**              | [**PolQA**](https://huggingface.co/datasets/ipipan/polqa)   | [**Allegro FAQ**](https://huggingface.co/datasets/piotr-rybak/allegro-faq) | [**Legal Questions**](https://huggingface.co/datasets/piotr-rybak/legal-questions) | **Average** |
+|:-----------------------|------------:|------------:|------------:|------------:|
+| BM25                   | 61.35       | 66.89       | **96.38**   | 74.87       |
+| BM25 (lemma)           | 71.49       | 75.33       | 94.57       | 80.46       |
+| MiniLM-L12-v2          | 37.24       | 71.67       | 78.97       | 62.62       |
+| LaBSE                  | 46.23       | 67.11       | 81.34       | 64.89       |
+| mContriever-Base       | 78.66       | 84.44       | 95.82       | 86.31       |
+| E5-Base                | 86.61       | 91.89       | 96.24       | 91.58       |
+| ST-DistilRoBERTa       | 48.43       | 84.89       | 88.02       | 73.78       |
+| ST-MPNet               | 56.80       | 86.00       | 87.19       | 76.66       |
+| HerBERT-QA             | 75.84       | 85.78       | 91.09       | 84.23       |
+| **SilverRetriever**    | **87.24**   | **94.56**   | 95.54       | **92.45**   |
+### NDCG@10
+| **Model**              | **PolQA** | **Allegro FAQ** | **Legal Questions** | **Average** |
+|:-----------------------|-------------:|-------------:|-------------:|-------------:|
+| BM25                   | 24.51        | 48.71        | **82.21**    | 51.81        |
+| BM25 (lemma)           | 31.97        | 55.70        | 78.65        | 55.44        |
+| MiniLM-L12-v2          | 11.93        | 51.25        | 54.44        | 39.21        |
+| LaBSE                  | 15.53        | 46.71        | 56.16        | 39.47        |
+| mContriever-Base       | 36.30        | 67.38        | 77.42        | 60.37        |
+| E5-Base                | **46.08**    | 75.90        | 77.69        | 66.56        |
+| ST-DistilRoBERTa       | 16.73        | 64.39        | 63.76        | 48.29        |
+| ST-MPNet               | 21.55        | 65.44        | 62.99        | 49.99        |
+| HerBERT-QA             | 32.52        | 63.58        | 66.99        | 54.36        |
+| **SilverRetriever**    | 43.40        | **79.66**    | 77.10        | **66.72**    |
+## Usage
+### Preparing inputs
 The model was trained on question-passage pairs and works best when the input is the same format as that used during training:
+- We added the phrase `Pytanie:` to the beginning of the question.
 - The training passages consisted of `title` and `text` concatenated with the special token `</s>`. Even if your passages don't have a `title`, it is still beneficial to prefix a passage with the `</s>` token.
 - Although we used the dot product during training, the model usually works better with the cosine distance.
+### Inference with Sentence-Transformers
 Using this model becomes easy when you have [sentence-transformers](https://www.SBERT.net) installed:
 print(embeddings)
 ```
+### Inference with HuggingFace Transformers
 Without [sentence-transformers](https://www.SBERT.net), you can use the model like this: First, you pass your input through the transformer model, then you have to apply the right pooling-operation on-top of the contextualized word embeddings.
 ```python