CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nomic-ai /vdr-multilingual-trainimage100K<n<1M0 likes466 downloads2y agoHugging Face02shanchen /aime_2025_multilingualWhen Models Reason in Your Language: Controlling Thinking Trace Language Comes at the Cost of Accuracy https://arxiv.org/abs/2505.22888 Jirui Qi, Shan Chen, Zidi Xiong, Raquel Fernández, Danielle S. Bitterman, Arianna Bisazza Recent Large Reasoning Models (LRMs) with thinking traces have shown strong performance on English reasoning tasks. However, their ability to think in other languages is less studied. This capability is as important as answer accuracy for real world applications because… See the full description on the dataset page: https://huggingface.co/datasets/shanchen/aime_2025_multilingual.tabularn<1K0 likes371 downloads1y agoHugging Face03ellamind /aime26-multilingualtextn<1K0 likes359 downloads7mo agoHugging Face04fedric95 /AIME2025-Multilingual Description This repository contains a multi language version of the AIME2025 dataset. As the english reference version, we haved used the one created by the authors of MathArena. For completness, we have included the english version also in this repository, please, refer to the one contained in the MathArena github repository for the original one (https://github.com/eth-sri/matharena/tree/main/data/aime). Many thanks to Jasper Dekoninck for the help in understanding the structure… See the full description on the dataset page: https://huggingface.co/datasets/fedric95/AIME2025-Multilingual.tabularn<1K3 likes340 downloads10mo agoHugging Face05nomic-ai /vdr-multilingual-train-corpusimage100K<n<1M0 likes333 downloads2y agoHugging Face06ellamind /aime25-multilingualtextn<1K0 likes228 downloads7mo agoHugging Face07projecte-aina /RAG_Multilingual Dataset Card for RAG_Multilingual Dataset Summary RAG_Multilingual is an instruction-following synthetic QA dataset created from extractive QA datasets from Catalan, English and Spanish reference sets. The reference datasets were: SQAD (https://huggingface.co/datasets/rajpurkar/squad), Catalanqa (https://huggingface.co/datasets/projecte-aina/catalanqa) and SQAC (https://huggingface.co/datasets/PlanTL-GOB-ES/SQAC). This dataset, of 56.406 instances, was created by… See the full description on the dataset page: https://huggingface.co/datasets/projecte-aina/RAG_Multilingual.textquestion-answering10K<n<100K23 likes176 downloads2y agoHugging Face08AIML-TUDA /LongBench-multilingualWIP, please don't use yet text10K<n<100K1 likes120 downloads7mo agoHugging Face09AI-Culture-Commons /ai-culture-multilingual-json-dolma AI-Culture Multilingual JSON + DOLMA Corpus 16M words · 12 languages · CC-BY-4.0 The AI-Culture corpus contains 5K articles providing comprehensive philosophical and cultural content, exploring the intersection of technology, artificial intelligence, and human culture, perfectly aligned across 12 languages. All content maintains identical parallel structure across translations with zero duplication and editor-curated quality. This project is maintained by a non-profit digital… See the full description on the dataset page: https://huggingface.co/datasets/AI-Culture-Commons/ai-culture-multilingual-json-dolma.texttranslation1K<n<10K3 likes92 downloads1y agoHugging Face10empero-ai /tasklist-grok4-multilingual-50000x-unfiltered TaskGen Dataset Generated with taskgen by empero-org Run Parameters Parameter Value Model grok-4-1-fast-non-reasoning Temperature 0.75 Total Tasks 50000 Concurrency 8 workers API Base https://api.x.ai/v1 Generated 2026-04-07 09:04:57 Budget Cap $15.0000 Multilingual Yes (en, de, fr, es, nl, zh, ar, ru) Language Distribution Language Code Tasks Arabic ar 6111 Chinese zh 6058 German de 6057 Spanish es 6020… See the full description on the dataset page: https://huggingface.co/datasets/empero-ai/tasklist-grok4-multilingual-50000x-unfiltered.tabular10K<n<100K2 likes81 downloads6mo agoHugging Face11AISE-TUDelft /multilingual-code-comments-fixed-8 Fixed-8 Based on fixed-7 revision 14e85fe00a8b284cd226c58281ddd8e6b990b190. Replaces six Greek rows with missing expert labels with six newly labelled samples. All five language configurations retain 500 training rows (2,500 total). All other rows are unchanged. Removed ID Replacement ID 8000_5 1056_0 8000_14 4357_8 8000_15 4848_9 8000_16 29069_13 8000_17 1385_4 8000_18 5142_0 All 500 Greek rows now have all five expert accuracy labels. Original… See the full description on the dataset page: https://huggingface.co/datasets/AISE-TUDelft/multilingual-code-comments-fixed-8.text1K<n<10K0 likes74 downloads7d agoHugging Face12appier-ai-research /multilingual-CulturalBench-Hardtabular10K<n<100K0 likes72 downloads2y agoHugging Face13empero-ai /tasklist-grok-multilingual-100000x-unfiltered TaskGen Dataset Generated with taskgen by empero-org Run Parameters Parameter Value Model grok-4-1-fast-reasoning Temperature 0.9 Total Tasks 83052 Concurrency 30 workers API Base https://api.x.ai/v1 Generated 2026-04-07 14:31:14 Budget Cap $15.0000 Multilingual Yes (en, de, fr, es, nl, zh, ar, ru) Language Distribution Language Code Tasks Arabic ar 10446 German de 10397 Dutch nl 10353 Spanish es 10345… See the full description on the dataset page: https://huggingface.co/datasets/empero-ai/tasklist-grok-multilingual-100000x-unfiltered.tabular100K<n<1M2 likes70 downloads6mo agoHugging Face14moonshine-ai /multilingual_examplesaudion<1K0 likes68 downloads1y agoHugging Face15wujoe132 /ponys-multilingual-ai-character-consistency-benchmark Ponys Multilingual AI Character Consistency Benchmark This repository contains a preregistered test instrument, not collected product results and not an independent product ranking. 140 fixed test cases across seven locales four dimensions: persona, register, relationship state, and visual identity three planned clean-session runs per case result state: not_collected publisher: Ponys.ai Research (official first-party research) official source: https://ponys.ai/ research feeds:… See the full description on the dataset page: https://huggingface.co/datasets/wujoe132/ponys-multilingual-ai-character-consistency-benchmark.tabulartext-generationn<1K0 likes65 downloads25d agoHugging Face16shanchen /aime_2024_multilingualWhen Models Reason in Your Language: Controlling Thinking Trace Language Comes at the Cost of Accuracy https://arxiv.org/abs/2505.22888 Jirui Qi, Shan Chen, Zidi Xiong, Raquel Fernández, Danielle S. Bitterman, Arianna Bisazza Recent Large Reasoning Models (LRMs) with thinking traces have shown strong performance on English reasoning tasks. However, their ability to think in other languages is less studied. This capability is as important as answer accuracy for real world applications because… See the full description on the dataset page: https://huggingface.co/datasets/shanchen/aime_2024_multilingual.textn<1K0 likes54 downloads1y agoHugging Face17lightonai /aime24_multilingual AIME24 Multilingual aime24_multilingual is a multilingual version of the benchmark AIME 2024, covering six languages: English, French, German, Spanish, Chinese, and Swahili. Each sample is a competition-level mathematics problem from the American Invitational Mathematics Examination (AIME) 2024, translated into the five target languages. This release is a corrected version of shanchen/aime_2024_multilingual that fixes translation artifacts and errors. It is released alongside the… See the full description on the dataset page: https://huggingface.co/datasets/lightonai/aime24_multilingual.textquestion-answeringn<1K0 likes53 downloads4mo agoHugging Face18AIML-TUDA /RULER-multilingualtabular10K<n<100K1 likes51 downloads6mo agoHugging Face19lightonai /aime25_multilingual AIME25 Multilingual aime25_multilingual is a multilingual version of the benchmark AIME 2025, covering six languages: English, French, German, Spanish, Chinese, and Swahili. Each sample is a competition-level mathematics problem from the American Invitational Mathematics Examination (AIME) 2025, translated into the five target languages. This release is a corrected version of shanchen/aime_2025_multilingual that fixes translation artifacts and errors. It is released alongside the… See the full description on the dataset page: https://huggingface.co/datasets/lightonai/aime25_multilingual.textquestion-answeringn<1K0 likes51 downloads4mo agoHugging Face20nomic-ai /vdr-multilingual-train-hn-minetext100K<n<1M0 likes49 downloads2y agoHugging Face21appier-ai-research /MATH-multilingual-traintext1K<n<10K0 likes48 downloads2y agoHugging Face22AIM-Harvard /cardiffnlp_tweet_sentiment_multilingual_translatedtext1K<n<10K0 likes45 downloads2y agoHugging Face23shanchen /aiw_easy_multilingualtext1K<n<10K0 likes42 downloads2y agoHugging Face24appier-ai-research /Multilingual-MATH-500text1K<n<10K1 likes39 downloads2y agoHugging Face25AIM-Harvard /multilingual_toxicity_datasettext10K<n<100K0 likes37 downloads2y agoHugging Face26shanchen /aiw_hard_multilingualtext1K<n<10K0 likes37 downloads2y agoHugging Face27appier-ai-research /multilingual-dilemmatext10K<n<100K0 likes36 downloads2y agoHugging Face28AISE-TUDelft /multilingual-code-comments A Qualitative Investigation into LLM-Generated Multilingual Code Comments and Automatic Evaluation Metrics This dataset helps us understand how Large Language Models (LLMs) can create code comments in different languages. While LLMs are good at coding tasks in English, we don't know much about how well they work in other languages. This dataset, along with our research, studies how LLMs generate code comments in English, Chinese, Dutch, Polish, and Greek. In our case, we have… See the full description on the dataset page: https://huggingface.co/datasets/AISE-TUDelft/multilingual-code-comments.text1K<n<10K3 likes36 downloads1y agoHugging Face29Metric-AI /open-asr-leaderboard-multilingual-datasets Open ASR Leaderboard Armenian Test Datasets This private repository holds leaderboard-compatible Armenian test configurations while their integration is being validated. Configurations fleurs_hy Source: google/fleurs, configuration hy_am, test split Reviewed reference changes: Metric-AI/fleurs-corrections, test split 932 recordings; all 314 reviewed corrections were matched to the original source transcript and applied mcv_hy… See the full description on the dataset page: https://huggingface.co/datasets/Metric-AI/open-asr-leaderboard-multilingual-datasets.audioautomatic-speech-recognition1K<n<10K1 likes35 downloads14d agoHugging Face30AISE-TUDelft /multilingual-code-comments-fixed-7text1K<n<10K0 likes33 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.