imvladikon/knesset_meetings_corpus
Dataset Card Dataset Summary An example of a sample: { "text": <text content of given document>, "path": <file path to docx> } Dataset usage Available "kneset16","kneset17","knesset_tagged" configurations And only train set. train_ds = load_dataset("imvladikon/knesset_meetings_corpus", "kneset16", split="train") The Knesset Meetings Corpus 2004-2005 is made up of two components: Raw texts - 282 files made up of 867,725 lines together. These can be… See the full description on the dataset page: https://huggingface.co/datasets/imvladikon/knesset_meetings_corpus.
Fix task_categories (#2)
Fix `license` metadata (#1)
Update README.md
Update README.md
Update README.md
Create README.md
init
Update README.md
initial commit
