CoolFace
Datasetpublic

multilingual-discourse-hub/disrpt

Disrpt is a multilingual, multi-framework unified discourse analysis benchmark. It unifies discourse relation classification tasks (.rels) and discourse segmentation (.connlu) for many languages. ⚠️ This repo only contains the disrpt dataset when the underlying data is permissively licensed. Some datasets rely on corpora like the PTB. To load these datasets, run the following: pip install disrpt-utils Then from disrpt_utils import load_dataset corpora_paths={ # ⚠️✍️ TODO… See the full description on the dataset page: https://huggingface.co/datasets/multilingual-discourse-hub/disrpt.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
3likes12kdownloads
settings

This repository belongs to multilingual-discourse-hub on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

namedisrpt
visibilitypublic
licenceapache-2.0
gatedno
ownermultilingual-discourse-hub
Account settings
multilingual-discourse-hub/disrpt · CoolFace