CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01m-hamza-mughal /beat2-additional-annotations BEAT2 Official Release + Additional Annotations This is a fork of H-Liu1997/BEAT2 that adds annotations contributed by the RAG-Gesture (CVPR 2025) and MIBURI (CVPR 2026) projects. The base BEAT2-English data (motion, audio, TextGrids, semantic labels, pretrained motion-autoencoder weights) is inherited verbatim from upstream; the additional annotations from RAG-Gesture and MIBURI are pushed on top. Citations If you use only the original BEAT2 dataset, please cite… See the full description on the dataset page: https://huggingface.co/datasets/m-hamza-mughal/beat2-additional-annotations.audio1K<n<10K0 likes2.3k downloads3mo agoHugging Face02kanhatakeyama /wizardlm8x22b-logical-math-coding-sft_additional 自動生成したテキスト WizardLM 8x22bで生成した論理・数学・コード系のデータです。 一部の計算には東京工業大学のスーパーコンピュータTSUBAME4.0を利用しました。 text100K<n<1M0 likes578 downloads2y agoHugging Face03mscho331 /bank-additional-fulltext10K<n<100K1 likes548 downloads1y agoHugging Face04mib-bench /arithmetic_additiontabular10K<n<100K0 likes397 downloads1y agoHugging Face05JakeOh /addition-datasettext1M<n<10M0 likes364 downloads10mo agoHugging Face06IEMaster /worldedit_addition_v2tabular10K<n<100K0 likes354 downloads2y agoHugging Face07deqing /addition_dataset Addition Dataset Addition problems in the format {a} + {b} = {c}. Subsets test: 5K held-out evaluation examples (operands >= 10, i.e. min 2 digits) 1BT: 85M training examples (1 billion tokens under Llama-3 tokenizer) 10BT: 850M training examples (10 billion tokens) 3MT-3digit: Exhaustive single-token addition: all (a, b) with a, b in [0, 999] and a+b <= 999. 500,500 ordered pairs, ~3M tokens. All of a, b, c are single tokens. Symmetry-safe train/test split (10% test).… See the full description on the dataset page: https://huggingface.co/datasets/deqing/addition_dataset.text1B<n<10B0 likes218 downloads4mo agoHugging Face08Lots-of-LoRAs /task753_svamp_addition_question_answering Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task753_svamp_addition_question_answering Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+ NLP… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task753_svamp_addition_question_answering.texttext-generationn<1K0 likes101 downloads2y agoHugging Face09sohampnow /slam_stage2_additional_datatextn<1K0 likes75 downloads2y agoHugging Face10marcov /openbookqa_additional_promptsourcetabular10K<n<100K0 likes68 downloads2y agoHugging Face11selfcorrexp /llama3_additional_rr40k_non_delete_sfttabular100K<n<1M0 likes66 downloads2y agoHugging Face12IEMaster /worldedit_addition_v1tabular10K<n<100K0 likes65 downloads2y agoHugging Face13DiCeyIII /Additional_Yoruba_Dataaudio1K<n<10K0 likes61 downloads2y agoHugging Face14garrethlee /bpe-single-multi-token-additiontext10K<n<100K0 likes61 downloads2y agoHugging Face15EleutherAI /quirky_addition_increment0 Dataset Card for "quirky_addition_increment0" More Information needed text100K<n<1M0 likes53 downloads3y agoHugging Face16selfcorrexp /llama3_additional_rr80k_NON_balanced_sfttabular100K<n<1M0 likes53 downloads2y agoHugging Face17EleutherAI /quirky_addition_rawtabular100K<n<1M0 likes46 downloads3y agoHugging Face18selfcorrexp2 /llama31_no_additional_chat_formattabular100K<n<1M0 likes46 downloads2y agoHugging Face19selfcorrexp /llama3_additional_rr40k_non_delete_sft_chat_formattabular100K<n<1M0 likes45 downloads2y agoHugging Face20atmallen /quirky_addition_increment3_bob_hard Dataset Card for "quirky_addition_increment3_bob_hard" More Information needed text10K<n<100K0 likes44 downloads3y agoHugging Face21flexitok /multilingual-addition Multilingual Addition Dataset Synthetic dataset of addition problems of the form a+b=answer, where a and b are written-form representations of integers in 21 languages, plus a 22nd split using raw digit strings. Task format Each sample contains: field type description a_str str written-form (or digit) representation of a a_digit int integer value of a b_str str written-form (or digit) representation of b b_digit int integer value of b answer str… See the full description on the dataset page: https://huggingface.co/datasets/flexitok/multilingual-addition.tabularquestion-answering10M<n<100M0 likes44 downloads5mo agoHugging Face22selfcorrexp /llama3_additional_rr40k_NON_balanced_sfttabular100K<n<1M0 likes43 downloads2y agoHugging Face23MichaelAnthony /lemonseed-addition-carry-method lemonseed-addition-carry-method LemonSeed — carry-method addition scratchpad (LSB-first, single-digit facts). Teaches digit-level addition with explicit written steps. Contents addition.jsonl (10000 rows) Format JSON Lines (.jsonl), one example per line. Provenance Synthetic, generated programmatically for the LemonSeed 1.5B project (by Geramy L. Loveless). Data authored by Michael Anthony Falabella. textquestion-answering10K<n<100K0 likes43 downloads28d agoHugging Face24selfcorrexp /llama3_additional_rr40k_balanced_sfttabular100K<n<1M0 likes42 downloads2y agoHugging Face25selfcorrexp /llama3_additional_rr10k_NON_balanced_sfttabular100K<n<1M0 likes41 downloads2y agoHugging Face26mkyle /addition_dataset_5_milliontabular1M<n<10M0 likes40 downloads10mo agoHugging Face27intagliated /gaperon-distill-additionstabular1K<n<10K0 likes40 downloads2mo agoHugging Face28EleutherAI /quirky_addition_increment0_alice Dataset Card for "quirky_addition_increment0_alice" More Information needed text100K<n<1M0 likes39 downloads3y agoHugging Face29clue2solve /langchain-additional-resourcestextn<1K0 likes37 downloads3y agoHugging Face30kanhatakeyama /wizardlm8x22b-logical-math-coding-sft_additional-ja 自動生成したテキスト WizardLM 8x22bで生成した論理・数学・コード系のデータを、Calm3-22bで翻訳したものです。 一部の計算には東京工業大学のスーパーコンピュータTSUBAME4.0を利用しました text10K<n<100K9 likes37 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.