CoolFace
Datasetpublic

multilingual-discourse-hub/disrpt

Disrpt is a multilingual, multi-framework unified discourse analysis benchmark. It unifies discourse relation classification tasks (.rels) and discourse segmentation (.connlu) for many languages. ⚠️ This repo only contains the disrpt dataset when the underlying data is permissively licensed. Some datasets rely on corpora like the PTB. To load these datasets, run the following: pip install disrpt-utils Then from disrpt_utils import load_dataset corpora_paths={ # ⚠️✍️ TODO… See the full description on the dataset page: https://huggingface.co/datasets/multilingual-discourse-hub/disrpt.

sourceHugging Faceapache-2.0updated 1y agoView on Hugging Face
3likes12kdownloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
multilingual-discourse-hub/disrpt · CoolFace