CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01vblagoje /lfqatext100K<n<1M16 likes197 downloads5y agoHugging Face02vblagoje /lfqa_support_docsSupport documents for building https://huggingface.co/vblagoje/bart_lfqa model text100K<n<1M6 likes162 downloads5y agoHugging Face03indonesian-nlp /lfqa_idtext100K<n<1M4 likes86 downloads5y agoHugging Face04harsha28 /legal-reasoning-lfqa-merged Dataset Card for "legal-reasoning-lfqa-merged" More Information needed text10K<n<100K2 likes86 downloads3y agoHugging Face05fahdsoliman /lfqa_test_docstext100K<n<1M0 likes75 downloads2y agoHugging Face06stefanbschneider /lfqa-max-answer-length-512 Dataset Card Dataset Description The dataset contains simple, long-form answers to questions and corresponding contexts. Similar to ELI5 but with context. This dataset is a filtered version of LLukas22/lfqa_preprocessed, which in turn is a processed and simplified version of of vblagoje's lfqa_support_docs and lfqa datasets. I have filtered out overly long answers, based on the number of tokens in the answer using the LED tokenizer. It can be reproduced with the notebook… See the full description on the dataset page: https://huggingface.co/datasets/stefanbschneider/lfqa-max-answer-length-512.textquestion-answering100K<n<1M1 likes72 downloads2y agoHugging Face07avacaondata /lfqa_squadtext100K<n<1M0 likes67 downloads5y agoHugging Face08harshasurampudi /legal-reasoning-lfqa-synthetic Dataset Card for "legal-reasoning-lfqa-synthetic" More Information needed text10K<n<100K6 likes50 downloads3y agoHugging Face09stefanbschneider /lfqa_preprocessed Dataset Card for "stefanbschneider/lfqa_preprocessed" Dataset Description The dataset contains simple, long-form answers to questions and corresponding contexts. Similar to ELI5 but with context. This dataset is a filtered version of LLukas22/lfqa_preprocessed, which in turn is a processed and simplified version of of vblagoje's lfqa_support_docs and lfqa datasets. This dataset (stefanbschneider/lfqa_preprocessed) has filtered out overly long answers (around 3%). It can… See the full description on the dataset page: https://huggingface.co/datasets/stefanbschneider/lfqa_preprocessed.textquestion-answering100K<n<1M0 likes48 downloads2y agoHugging Face10stefanbschneider /lfqa-max-answer-length-1024 Dataset Card Dataset Description The dataset contains simple, long-form answers to questions and corresponding contexts. Similar to ELI5 but with context. This dataset is a filtered version of LLukas22/lfqa_preprocessed, which in turn is a processed and simplified version of of vblagoje's lfqa_support_docs and lfqa datasets. I have filtered out overly long answers, based on the number of tokens in the answer using the LED tokenizer. It can be reproduced with the notebook… See the full description on the dataset page: https://huggingface.co/datasets/stefanbschneider/lfqa-max-answer-length-1024.textquestion-answering100K<n<1M0 likes47 downloads2y agoHugging Face11efederici /lfqa-preprocessed-ittextquestion-answering10K<n<100K2 likes45 downloads3y agoHugging Face12adamjweintraut /eli5_lfqa_besttabular100K<n<1M1 likes43 downloads3y agoHugging Face13fangyuan /lfqa_discourseLFQA discourse contains discourse annotations of long-form answers. - [VALIDITY]: Validity annotations of (question, answer) pairs. - [ROLE]: Role annotations of valid answer paragraphs.1K<n<10K1 likes38 downloads3y agoHugging Face14LLukas22 /lfqa_preprocessed Dataset Card for "lfqa_preprocessed" Dataset Summary This is a simplified version of vblagoje's lfqa_support_docs and lfqa datasets. It was generated by me to have a more straight forward way to train Seq2Seq models on context based long form question answering tasks. Dataset Structure Data Instances An example of 'train' looks as follows. { "question": "what's the difference between a forest and a wood?", "answer": "They're used… See the full description on the dataset page: https://huggingface.co/datasets/LLukas22/lfqa_preprocessed.textquestion-answering100K<n<1M2 likes36 downloads4y agoHugging Face15aitetic /eli5-lfqa ELI5: Long Form Question Answering https://arxiv.org/pdf/1907.09190 Items (train): 272_634 Items (val): 1_507 Downloaded from: https://www.kaggle.com/datasets/trandaiphu/eli5-dataset Tokens ctxs length distribution summary: min: 152 mean: 3745.35 median (p50): 3736.00 p90: 4037.00 p95: 4152.00 p99: 4432.00 max: 6530 Tokens answers length distribution summary: rows: 272_634 min: 8 mean: 401.31 median (p50): 217.00 p90: 820.70 p95: 1230.00 p99:… See the full description on the dataset page: https://huggingface.co/datasets/aitetic/eli5-lfqa.text100K<n<1M0 likes33 downloads2mo agoHugging Face16NinaCalvi /lfqa_expert_pairwise_human_preference_no_reasoningtabularn<1K0 likes32 downloads2y agoHugging Face17nlpatunt /LFQA-HP-1M LFQA-HP-1M: A Large-Scale Human Preference Dataset for Long-Form Question Answering Overview LFQA-HP-1M is a large-scale human preference dataset for Long-Form Question Answering (LFQA). The dataset is designed to support research in: Human preference modeling Pairwise answer evaluation Fine-grained rubric-based scoring LLM-as-a-Judge evaluation Long-form response generation benchmarking LFQA tasks require multi-sentence, explanatory, reasoning-based answers rather than… See the full description on the dataset page: https://huggingface.co/datasets/nlpatunt/LFQA-HP-1M.question-answering1M<n<10M1 likes32 downloads7mo agoHugging Face18ericholam /lfqatext10K<n<100K0 likes32 downloads11mo agoHugging Face19fahdsoliman /lfqa_with_supports_subsettext1K<n<10K1 likes31 downloads2y agoHugging Face20hanane /attributionBench_lfqa_expertqa_dpo Dataset Card for "attributionBench_lfqa_expertqa_dpo" More Information needed tabular1K<n<10K0 likes25 downloads1y agoHugging Face21allenai /intent-aware-lfqa-intent-implicittext1K<n<10K1 likes20 downloads6mo agoHugging Face22adamjweintraut /bart-finetuned-eli5_lfqa_best_slice-512_2023-12-10_runtabular1K<n<10K0 likes19 downloads3y agoHugging Face23ContextualAI /LFQAtabularn<1K0 likes19 downloads1y agoHugging Face24adamjweintraut /eli5_lfqa_best_slicetabular10K<n<100K1 likes17 downloads3y agoHugging Face25allenai /intent-aware-lfqa-multiviewtext1K<n<10K1 likes14 downloads6mo agoHugging Face26aitetic /eli5-lfqa-combined eli5-lfqa-combined Postprocessed from eli5-lfqa: (ctxs, question+answers[]) Assembled by: dataset_assembler.py Size: 1.1B text100K<n<1M0 likes14 downloads2mo agoHugging Face27allenai /intent-aware-lfqa-baselinetext1K<n<10K1 likes13 downloads6mo agoHugging Face28voidful /lfqa_eval Dataset Card for "lfqa_eval" More Information needed text1K<n<10K0 likes12 downloads3y agoHugging Face29ContextualAI /LFQA_eval_dataset_unit_tests_justificationtabularn<1K0 likes12 downloads2y agoHugging Face30nlpatunt /lfqa-textbuggertextn<1K0 likes12 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.