CoolFace
19 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01embedded-language-flows /xsum_train_t5tabular100K<n<1M0 likes400 downloads5mo agoHugging Face02Gabriel /xsum_swe Dataset Card for Swedish Xsum Dataset The Swedish xsum dataset has only been machine-translated to improve downstream fine-tuning on Swedish summarization tasks. Dataset Summary Read about the full details at original English version: https://huggingface.co/datasets/xsum Data Fields id: a string containing the heximal formated SHA1 hash of the url where the story was retrieved from document: a string containing the body of the news article summary: a string… See the full description on the dataset page: https://huggingface.co/datasets/Gabriel/xsum_swe.tabularsummarization100K<n<1M0 likes145 downloads4y agoHugging Face03stacked-summaries /stacked-xsum xsum-stacked The current version (corresponding to the stacked-booksum release): v0.3. See the Stacked Summaries org page for what this is and why it exists. The maximum input length is 16384 tokens, and the maximum output length is 1024 tokens (measured with the Long-T5 tokenizer). stats [2023-01-09 19:36:25] INFO:root:INPUTS - basic stats - train [2023-01-09 19:36:26] INFO:root:{'num_columns': 5, 'num_rows': 204045, 'num_unique_target': 203107, 'num_unique_text':… See the full description on the dataset page: https://huggingface.co/datasets/stacked-summaries/stacked-xsum.tabularsummarization100K<n<1M2 likes128 downloads4y agoHugging Face04stacked-summaries /stacked-xsum-1024 stacked-xsum-1024 a "stacked" version of xsum Original Dataset: copy of the base dataset Stacked Rows: The original dataset is processed by stacking rows based on certain criteria: Maximum Input Length: The maximum length for input sequences is 1024 tokens in the longt5 model tokenizer. Maximum Output Length: The maximum length for output sequences is also 1024 tokens in the longt5 model tokenizer. Special Token: The dataset utilizes the [NEXT_CONCEPT] token to indicate a new… See the full description on the dataset page: https://huggingface.co/datasets/stacked-summaries/stacked-xsum-1024.tabularsummarization100K<n<1M1 likes94 downloads3y agoHugging Face05bobox /xSum-processedtabular100K<n<1M0 likes81 downloads2y agoHugging Face06stacked-summaries /onlystacked-xsum-1024 stacked-summaries/onlystacked-xsum-1024 Same thing as stacked-summaries/stacked-xsum-1024 but filtered such that is_stacked=True. Please refer to the original dataset for info and to raise issues if needed. Basic info on train split: <class 'pandas.core.frame.DataFrame'> RangeIndex: 116994 entries, 0 to 116993 Data columns (total 6 columns): # Column Non-Null Count Dtype --- ------ -------------- ----- 0 document 116994 non-null string 1… See the full description on the dataset page: https://huggingface.co/datasets/stacked-summaries/onlystacked-xsum-1024.tabularsummarization100K<n<1M0 likes70 downloads3y agoHugging Face07whu9 /xsum_postprocess Dataset Card for "xsum_postprocess" More Information needed tabular100K<n<1M0 likes41 downloads3y agoHugging Face08rubricreward /R3-eval-XSUMtabular1K<n<10K0 likes32 downloads1y agoHugging Face09fabhiansan /XSum-Indonesia-with-Entailment-Labeltabular10K<n<100K0 likes29 downloads1y agoHugging Face10fabhiansan /XSum-Indonesia-with-Perturbationtabular100K<n<1M0 likes22 downloads1y agoHugging Face11linluqiu /xsum_train_target_64_context_1024tabular100K<n<1M0 likes22 downloads8mo agoHugging Face12fabhiansan /XSUM-Indonesia-AMR-NLI XSUM-Indonesia-AMR-NLI Deskripsi Dataset ini berisi kumpulan data dalam bahasa Indonesia yang dirancang untuk tugas Natural Language Inference (NLI). Setiap instans data terdiri dari teks sumber (source_text) yang ada pada dataset XSum teks yang dihasilkan (generated_indonesian yang berfungsi sebagai hipotesis dibuat dari AMR Perturbasi) skor (score) yang menunjukkan hubungan antara keduanya (0 untuk non-entailment, 1 untuk entailment). ringkasan asli (target_summary)… See the full description on the dataset page: https://huggingface.co/datasets/fabhiansan/XSUM-Indonesia-AMR-NLI.tabular10K<n<100K0 likes20 downloads1y agoHugging Face13mtc /cleaned_xsum-faith-test-set-with-faithfulness-annotation Dataset Card for "cleaned_xsum-faith-test-set-with-faithfulness-annotation" More Information needed tabular1K<n<10K0 likes13 downloads3y agoHugging Face14fabhiansan /XSum-Indonesia-Entails-Onlytabular10K<n<100K0 likes13 downloads1y agoHugging Face15mtc /faithfulness_benchmark_sanity_check_xsum_faith Dataset Card for "faithfulness_benchmark_sanity_check_xsum_faith" More Information needed tabularn<1K0 likes10 downloads3y agoHugging Face16P-I2 /RESULTS_LLAMA13b_SPV_MIA_XSUM_evaltabular1K<n<10K0 likes7 downloads2y agoHugging Face17TheFactoryX /edition_1448_stacked-summaries-stacked-xsum-readymade edition_1448_stacked-summaries-stacked-xsum-readymade A Readymade by TheFactoryX Original Dataset stacked-summaries/stacked-xsum Process This dataset is a "readymade" - inspired by Marcel Duchamp's concept of taking everyday objects and recontextualizing them as art. What we did: Selected the original dataset from Hugging Face Shuffled each column independently Destroyed all row-wise relationships Preserved structure, removed meaning The result: Same data.… See the full description on the dataset page: https://huggingface.co/datasets/TheFactoryX/edition_1448_stacked-summaries-stacked-xsum-readymade.tabularn<1K0 likes7 downloads9mo agoHugging Face18TheFactoryX /edition_1449_stacked-summaries-stacked-xsum-readymade edition_1449_stacked-summaries-stacked-xsum-readymade A Readymade by TheFactoryX Original Dataset stacked-summaries/stacked-xsum Process This dataset is a "readymade" - inspired by Marcel Duchamp's concept of taking everyday objects and recontextualizing them as art. What we did: Selected the original dataset from Hugging Face Shuffled each column independently Destroyed all row-wise relationships Preserved structure, removed meaning The result: Same data.… See the full description on the dataset page: https://huggingface.co/datasets/TheFactoryX/edition_1449_stacked-summaries-stacked-xsum-readymade.tabularn<1K0 likes4 downloads9mo agoHugging Face19TheFactoryX /edition_1436_stacked-summaries-stacked-xsum-readymade edition_1436_stacked-summaries-stacked-xsum-readymade A Readymade by TheFactoryX Original Dataset stacked-summaries/stacked-xsum Process This dataset is a "readymade" - inspired by Marcel Duchamp's concept of taking everyday objects and recontextualizing them as art. What we did: Selected the original dataset from Hugging Face Shuffled each column independently Destroyed all row-wise relationships Preserved structure, removed meaning The result: Same data.… See the full description on the dataset page: https://huggingface.co/datasets/TheFactoryX/edition_1436_stacked-summaries-stacked-xsum-readymade.tabularn<1K0 likes3 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.