CoolFace
25 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01interlinguistic-language-modeling /ilm_detext1M<n<10M0 likes54 downloads6mo agoHugging Face02xinyuzhou2000 /Towards-Joint-Modeling-of-Dialogue-Response-and-Speech-Synthesis-based-on-Large-Language-Modeltext10K<n<100K3 likes48 downloads3y agoHugging Face03interlinguistic-language-modeling /ilm_estext1M<n<10M0 likes40 downloads6mo agoHugging Face04interlinguistic-language-modeling /ilm_itatext1M<n<10M0 likes39 downloads4mo agoHugging Face05interlinguistic-language-modeling /ilm_poltext1M<n<10M0 likes31 downloads4mo agoHugging Face06interlinguistic-language-modeling /ilm_eutext1M<n<10M0 likes22 downloads6mo agoHugging Face07interlinguistic-language-modeling /ilm_sltext100K<n<1M0 likes19 downloads6mo agoHugging Face08patrikgerard /ru_language_modeling_v4text1K<n<10K0 likes18 downloads2y agoHugging Face09interlinguistic-language-modeling /ilm_nltext100K<n<1M0 likes17 downloads6mo agoHugging Face10langtech-languagemodeling /siqa_ca_old Dataset Card: SIQA_CA (Pre-revision version) Description SIQA_CA (Pre-revision) is an earlier Catalan translation of the Social IQa (SIQA) dataset, a benchmark designed to evaluate commonsense reasoning about social interactions. This version consists of manually translated instances from the original English dataset into Catalan. It is used as a baseline for comparison against a revised and improved version of the dataset (SIQA_CA v2). Motivation and Use Case… See the full description on the dataset page: https://huggingface.co/datasets/langtech-languagemodeling/siqa_ca_old.textquestion-answering1K<n<10K0 likes16 downloads5mo agoHugging Face11langtech-languagemodeling /aliaboost_when2call_estextn<1K0 likes15 downloads6mo agoHugging Face12interlinguistic-language-modeling /ilm_pltext1M<n<10M0 likes15 downloads6mo agoHugging Face13niwang66 /mobile-actions-language-modeling Mobile Actions SFT Dataset A converted version of the google/mobile-actions dataset for supervised fine-tuning (SFT) of Qwen models with tool calling capabilities. Dataset Description This dataset is derived from the google/mobile-actions dataset, which contains human-AI conversations about performing actions on mobile devices. The original dataset has been converted to the Qwen chat template format for efficient training of Qwen models. Conversion Process The… See the full description on the dataset page: https://huggingface.co/datasets/niwang66/mobile-actions-language-modeling.text1K<n<10K0 likes13 downloads6mo agoHugging Face14langtech-languagemodeling /ALIABOOST-C2textn<1K0 likes11 downloads7mo agoHugging Face15langtech-languagemodeling /summarization_gl_subsampledtext1K<n<10K0 likes9 downloads7mo agoHugging Face16patrikgerard /ru_language_modelingtext100K<n<1M0 likes8 downloads2y agoHugging Face17langtech-languagemodeling /aliaboost_format_following_estextn<1K0 likes8 downloads6mo agoHugging Face18patrikgerard /uk_language_modeling0 likes6 downloads2y agoHugging Face19patrikgerard /ru_language_modeling_v2text10M<n<100M0 likes6 downloads2y agoHugging Face20interlinguistic-language-modeling /ilm_ittext100K<n<1M0 likes6 downloads5mo agoHugging Face21patrikgerard /uk_language_modeling_v2text10M<n<100M0 likes5 downloads2y agoHugging Face22ravikumar1478 /masked_language_modeling_for_Telugu_languagetext10K<n<100K0 likes4 downloads3y agoHugging Face23TornikeShatbera /Corpus-for-Language-Modelingtext10K<n<100K0 likes1 downloads1y agoHugging Face24Tia25 /corpus_for_language_modelingtextn<1K0 likes1 downloads1y agoHugging Face25langtech-languagemodeling /piqa_es Dataset Card for PIQA (Spanish Version) Dataset summary This dataset provides the Spanish translation and adaptation of the validation set of PIQA (Physical Interaction: Question Answering). The original dataset was designed to evaluate physical commonsense reasoning in language models through questions about everyday situations. Each example presents a physical goal and two possible solutions, only one of which is correct. This Spanish adaptation enables… See the full description on the dataset page: https://huggingface.co/datasets/langtech-languagemodeling/piqa_es.textquestion-answering1K<n<10K0 likes9h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.