CoolFace
Datasetpublic

mbshr/XSUMUrdu-DW_BBC

Urdu_DW-BBC-512 Dataset Summary Urdu Summarization Dataset containining 76,637 records of Article + Summary pairs scrapped from BBC Urdu and DW Urdu News Websites. Preprocessed Version: upto 512 tokens (~words); removed URLs, Pic Captions etc Supported Tasks and Leaderboards Summarization: Extractive and Abstractive urT5 adapted from mT5 having monolingual vocabulary only; 40k tokens of Urdu. Fine-tuned version @… See the full description on the dataset page: https://huggingface.co/datasets/mbshr/XSUMUrdu-DW_BBC.

sourceHugging Facecc-by-4.0updated 2y agoView on Hugging Face
0likes24downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
mbshr/XSUMUrdu-DW_BBC · CoolFace