datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
BSARDRetrieval
BSARDRetrieval
An MTEB dataset
Massive Text Embedding Benchmark
The Belgian Statutory Article Retrieval Dataset (BSARD) is a French native dataset for studying legal information retrieval. BSARD consists of more than 22,600 statutory articles from Belgian law and about 1,100 legal questions posed by Belgian citizens and labeled by experienced jurists with relevant articles from the corpus.
Task category
t2t
Domains
Legal, Spoken
Reference… See the full description on the dataset page: https://huggingface.co/datasets/mteb/BSARDRetrieval.BSARDRetrieval.v2
BSARDRetrieval.v2
An MTEB dataset
Massive Text Embedding Benchmark
BSARD is a French native dataset for legal information retrieval. BSARDRetrieval.v2 covers multi-article queries, fixing issues (#2906) with the previous data loading.
Task category
t2c
Domains
Legal, Spoken
Reference
https://huggingface.co/datasets/maastrichtlawtech/bsard
Source datasets:
maastrichtlawtech/bsard
How to evaluate on this task
You can evaluate an embedding model on this… See the full description on the dataset page: https://huggingface.co/datasets/mteb/BSARDRetrieval.v2.multiple_choice_bsardbsagithub-issues
Dataset Card for "github-issues"
More Information needed
