Skip to content

Commit 3fdf27c

Browse files
authored
Merge pull request #1245 from hanhainebula/master
release training data for bge-multilingual-gemma2
2 parents 2efb004 + a547347 commit 3fdf27c

1 file changed

Lines changed: 1 addition & 0 deletions

File tree

dataset/README.md

Lines changed: 1 addition & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -8,6 +8,7 @@ This will point to the training data we use for training various models.
88
| [bge-m3-data](https://huggingface.co/datasets/Shitao/bge-m3-data) | Fine-tuning data used by [bge-m3](https://huggingface.co/BAAI/bge-m3) |
99
| [public-data](https://huggingface.co/datasets/cfli/bge-e5data) | Public data identical to [e5-mistral](https://huggingface.co/intfloat/e5-mistral-7b-instruct) |
1010
| [full-data](https://huggingface.co/datasets/cfli/bge-full-data) | The full dataset we used for training [bge-en-icl](https://huggingface.co/BAAI/bge-en-icl) |
11+
| [bge-multilingual-gemma2-data](https://huggingface.co/datasets/hanhainebula/bge-multilingual-gemma2-data) | The full multilingual dataset we used for training [bge-multilingual-gemma2](https://huggingface.co/BAAI/bge-multilingual-gemma2) |
1112
| [reranker-data](Shitao/bge-reranker-data) | a mixture of multilingual datasets |
1213

1314

0 commit comments

Comments
 (0)