CoolFace
Datasetpublic

samchain/BIS_speeches_97_23_MLM

Dataset Card for "BIS_Speeches_97_23" This dataset is built from scrapped speeches on the Bank of International Settlements thanks to this repo : https://github.com/HanssonMagnus/scrape_bis. The dataset is made of 12k speeches from 1997 to 2023. Each pair is built with extracted sentences from speeches, if B is following A then the 'next_sentence_label' is 1 else it is 0. Negative pairs are built by choosing a sentence from another speech randomly. Credits Full… See the full description on the dataset page: https://huggingface.co/datasets/samchain/BIS_speeches_97_23_MLM.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
2likes147downloads
9 commits on main
fb15f062y ago

Update README.md

samchain
c4a56852y ago

Update README.md

samchain
13a23163y ago

Update README.md

samchain
8edd5973y ago

Update README.md

samchain
e8311613y ago

Upload README.md with huggingface_hub

samchain
83454b03y ago

Upload data/test-00000-of-00001-961eb7f4f5d68fb7.parquet with huggingface_hub

samchain
6a8e7013y ago

Upload data/train-00001-of-00002-325742d49634d52b.parquet with huggingface_hub

samchain
6c820c93y ago

Upload data/train-00000-of-00002-6f005ccc8c807e0a.parquet with huggingface_hub

samchain
73b68fe3y ago

initial commit

samchain