CoolFace
Datasetpublic

imvladikon/knesset_meetings_corpus

Dataset Card Dataset Summary An example of a sample: { "text": <text content of given document>, "path": <file path to docx> } Dataset usage Available "kneset16","kneset17","knesset_tagged" configurations And only train set. train_ds = load_dataset("imvladikon/knesset_meetings_corpus", "kneset16", split="train") The Knesset Meetings Corpus 2004-2005 is made up of two components: Raw texts - 282 files made up of 867,725 lines together. These can be… See the full description on the dataset page: https://huggingface.co/datasets/imvladikon/knesset_meetings_corpus.

sourceHugging Facepddlupdated 4y agoView on Hugging Face
1likes41downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
imvladikon/knesset_meetings_corpus · CoolFace