CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ElliotJenkins /MLLM_pathtabularn<1K0 likes381 downloads1y agoHugging Face02Pawlo77 /mllm-shap MLLM-SHAP experiment datasets Curated test splits for studying Shapley-value explanations in multimodal large language models (text and audio inputs). Each configuration is a filtered, size-controlled subset built for reproducible benchmarking—not a full copy of the upstream corpora. Configs follow the naming pattern {task}__{source} (for example single_sentence__voice_bench). Quick load Pin a dataset revision for reproducibility (replace REVISION with the commit hash… See the full description on the dataset page: https://huggingface.co/datasets/Pawlo77/mllm-shap.tabulartext-generation1K<n<10K2 likes197 downloads4mo agoHugging Face03Logics-MLLM /Logics-SWE-Env-2.5K Logics-SWE-Env-2.5K 2,553 software engineering task instances · 1,771 repositories · 4 programming languages 🤗 Related model: Logics-SWE-Qwen3.6-27B 📄 Paper: One to More, More to One Overview What is this dataset? Logics-SWE-Env-2.5K is a collection of repository-level software engineering tasks for research on coding agents and environment-based reinforcement learning. It contains 2,553 unique task instances from 1,771 GitHub repositories… See the full description on the dataset page: https://huggingface.co/datasets/Logics-MLLM/Logics-SWE-Env-2.5K.tabular1K<n<10K2 likes103 downloads9h agoHugging Face04mvishiu11 /mllm-shap-new MLLM Shap Experiments Datasets This repo contains following datasets used in modality and multilinguality experiments using Shapley Values, including: Datasets based on Infinity Instruct Dataset: Multi Lingual Dataset - 99 rows (33 in english, 33 in french, 33 in spanish), all of them translated to remaining 2 languages - resulting in total of 297 rows. Datasets based on Voice Bench Dataset Multi Sentence Dataset - 250 rows in English, multi-sentence entries in english. Single… See the full description on the dataset page: https://huggingface.co/datasets/mvishiu11/mllm-shap-new.tabularn<1K0 likes21 downloads10mo agoHugging Face05mvishiu11 /mllm-shap-copy MLLM Shap Experiments Datasets This repo contains following datasets used in modality and multilinguality experiments using Shapley Values, including: Datasets based on Infinity Instruct Dataset: Multi Lingual Dataset - 99 rows (33 in english, 33 in french, 33 in spanish), all of them translated to remaining 2 languages - resulting in total of 297 rows. Datasets based on Voice Bench Dataset Multi Sentence Dataset - 250 rows in English, multi-sentence entries in english. Single… See the full description on the dataset page: https://huggingface.co/datasets/mvishiu11/mllm-shap-copy.tabularn<1K0 likes16 downloads7mo agoHugging Face06elichen-skymizer /mllm-ppl-evaluation-e2e-ground-truthtabularn<1K0 likes6 downloads5mo agoHugging Face07Salieri2077 /MLLM_routertabular1K<n<10K0 likes3 downloads10mo agoHugging Face08Salieri2077 /LlaVA-judge-MLLM_Routertabular1K<n<10K0 likes3 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.