CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01vblagoje /lfqatext100K<n<1M16 likes201 downloads5y agoHugging Face02vblagoje /lfqa_support_docsSupport documents for building https://huggingface.co/vblagoje/bart_lfqa model text100K<n<1M6 likes156 downloads5y agoHugging Face03indonesian-nlp /lfqa_idtext100K<n<1M4 likes93 downloads5y agoHugging Face04harsha28 /legal-reasoning-lfqa-merged Dataset Card for "legal-reasoning-lfqa-merged" More Information needed text10K<n<100K2 likes85 downloads3y agoHugging Face05fahdsoliman /lfqa_test_docstext100K<n<1M0 likes73 downloads2y agoHugging Face06stefanbschneider /lfqa-max-answer-length-512 Dataset Card Dataset Description The dataset contains simple, long-form answers to questions and corresponding contexts. Similar to ELI5 but with context. This dataset is a filtered version of LLukas22/lfqa_preprocessed, which in turn is a processed and simplified version of of vblagoje's lfqa_support_docs and lfqa datasets. I have filtered out overly long answers, based on the number of tokens in the answer using the LED tokenizer. It can be reproduced with the notebook… See the full description on the dataset page: https://huggingface.co/datasets/stefanbschneider/lfqa-max-answer-length-512.textquestion-answering100K<n<1M1 likes64 downloads2y agoHugging Face07avacaondata /lfqa_squadtext100K<n<1M0 likes62 downloads4y agoHugging Face08harshasurampudi /legal-reasoning-lfqa-synthetic Dataset Card for "legal-reasoning-lfqa-synthetic" More Information needed text10K<n<100K6 likes46 downloads3y agoHugging Face09efederici /lfqa-preprocessed-ittextquestion-answering10K<n<100K2 likes45 downloads3y agoHugging Face10stefanbschneider /lfqa-max-answer-length-1024 Dataset Card Dataset Description The dataset contains simple, long-form answers to questions and corresponding contexts. Similar to ELI5 but with context. This dataset is a filtered version of LLukas22/lfqa_preprocessed, which in turn is a processed and simplified version of of vblagoje's lfqa_support_docs and lfqa datasets. I have filtered out overly long answers, based on the number of tokens in the answer using the LED tokenizer. It can be reproduced with the notebook… See the full description on the dataset page: https://huggingface.co/datasets/stefanbschneider/lfqa-max-answer-length-1024.textquestion-answering100K<n<1M0 likes43 downloads2y agoHugging Face11stefanbschneider /lfqa_preprocessed Dataset Card for "stefanbschneider/lfqa_preprocessed" Dataset Description The dataset contains simple, long-form answers to questions and corresponding contexts. Similar to ELI5 but with context. This dataset is a filtered version of LLukas22/lfqa_preprocessed, which in turn is a processed and simplified version of of vblagoje's lfqa_support_docs and lfqa datasets. This dataset (stefanbschneider/lfqa_preprocessed) has filtered out overly long answers (around 3%). It can… See the full description on the dataset page: https://huggingface.co/datasets/stefanbschneider/lfqa_preprocessed.textquestion-answering100K<n<1M0 likes42 downloads2y agoHugging Face12LLukas22 /lfqa_preprocessed Dataset Card for "lfqa_preprocessed" Dataset Summary This is a simplified version of vblagoje's lfqa_support_docs and lfqa datasets. It was generated by me to have a more straight forward way to train Seq2Seq models on context based long form question answering tasks. Dataset Structure Data Instances An example of 'train' looks as follows. { "question": "what's the difference between a forest and a wood?", "answer": "They're used… See the full description on the dataset page: https://huggingface.co/datasets/LLukas22/lfqa_preprocessed.textquestion-answering100K<n<1M2 likes38 downloads4y agoHugging Face13adamjweintraut /eli5_lfqa_besttabular100K<n<1M1 likes37 downloads3y agoHugging Face14aitetic /eli5-lfqa ELI5: Long Form Question Answering https://arxiv.org/pdf/1907.09190 Items (train): 272_634 Items (val): 1_507 Downloaded from: https://www.kaggle.com/datasets/trandaiphu/eli5-dataset Tokens ctxs length distribution summary: min: 152 mean: 3745.35 median (p50): 3736.00 p90: 4037.00 p95: 4152.00 p99: 4432.00 max: 6530 Tokens answers length distribution summary: rows: 272_634 min: 8 mean: 401.31 median (p50): 217.00 p90: 820.70 p95: 1230.00 p99:… See the full description on the dataset page: https://huggingface.co/datasets/aitetic/eli5-lfqa.text100K<n<1M0 likes35 downloads2mo agoHugging Face15ericholam /lfqatext10K<n<100K0 likes33 downloads11mo agoHugging Face16NinaCalvi /lfqa_expert_pairwise_human_preference_no_reasoningtabularn<1K0 likes28 downloads2y agoHugging Face17fahdsoliman /lfqa_with_supports_subsettext1K<n<10K1 likes27 downloads2y agoHugging Face18ContextualAI /LFQAtabularn<1K0 likes21 downloads1y agoHugging Face19allenai /intent-aware-lfqa-intent-implicittext1K<n<10K1 likes20 downloads6mo agoHugging Face20adamjweintraut /bart-finetuned-eli5_lfqa_best_slice-512_2023-12-10_runtabular1K<n<10K0 likes19 downloads3y agoHugging Face21hanane /attributionBench_lfqa_expertqa_dpo Dataset Card for "attributionBench_lfqa_expertqa_dpo" More Information needed tabular1K<n<10K0 likes19 downloads1y agoHugging Face22aitetic /eli5-lfqa-combined eli5-lfqa-combined Postprocessed from eli5-lfqa: (ctxs, question+answers[]) Assembled by: dataset_assembler.py Size: 1.1B text100K<n<1M0 likes18 downloads2mo agoHugging Face23adamjweintraut /eli5_lfqa_best_slicetabular10K<n<100K1 likes17 downloads3y agoHugging Face24nlpatunt /lfqa-textbuggertextn<1K0 likes16 downloads7mo agoHugging Face25voidful /lfqa_eval Dataset Card for "lfqa_eval" More Information needed text1K<n<10K0 likes13 downloads3y agoHugging Face26ContextualAI /LFQA_eval_dataset_unit_tests_justificationtabularn<1K0 likes13 downloads2y agoHugging Face27allenai /intent-aware-lfqa-multiviewtext1K<n<10K1 likes13 downloads6mo agoHugging Face28SIA86 /LFQAKnowledgeBasetextquestion-answeringn<1K0 likes12 downloads3y agoHugging Face29allenai /intent-aware-lfqa-baselinetext1K<n<10K1 likes12 downloads6mo agoHugging Face30adamjweintraut /eli5_lfqa_toptabular100K<n<1M0 likes11 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.