CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01qimma /leaderboard-detailstext1M<n<10M0 likes24k downloads29d agoHugging Face02open-llm-leaderboard-old /details_tiiuae__falcon-180B Dataset Card for Evaluation run of tiiuae/falcon-180B Dataset Summary Dataset automatically created during the evaluation run of model tiiuae/falcon-180B on the Open LLM Leaderboard. The dataset is composed of 66 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 32 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_tiiuae__falcon-180B.1 likes4.9k downloads3y agoHugging Face03SaylorTwift /details_meta-llama__Llama-3.1-8B-Instruct_private Dataset Card for Evaluation run of meta-llama/Llama-3.1-8B-Instruct Dataset automatically created during the evaluation run of model meta-llama/Llama-3.1-8B-Instruct. The dataset is composed of 78 configuration, each one corresponding to one of the evaluated task. The dataset has been created from 20 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/SaylorTwift/details_meta-llama__Llama-3.1-8B-Instruct_private.textn<1K0 likes4.5k downloads1y agoHugging Face04mitermix /audiosnippets_small_with_detailed_annotationaudio100K<n<1M1 likes2.8k downloads2y agoHugging Face05mitermix /audiosnippets_small_with_detailed_annotation2audio1M<n<10M1 likes2.7k downloads2y agoHugging Face06OALL /details_CohereForAI__c4ai-command-r7b-arabic-02-2025_v2 Dataset Card for Evaluation run of CohereForAI/c4ai-command-r7b-arabic-02-2025 Dataset automatically created during the evaluation run of model CohereForAI/c4ai-command-r7b-arabic-02-2025. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_CohereForAI__c4ai-command-r7b-arabic-02-2025_v2.text100K<n<1M0 likes2.5k downloads2y agoHugging Face07open-llm-leaderboard-old /details_meta-llama__Llama-2-7b-hf Dataset Card for Evaluation run of meta-llama/Llama-2-7b-hf Dataset Summary Dataset automatically created during the evaluation run of model meta-llama/Llama-2-7b-hf on the Open LLM Leaderboard. The dataset is composed of 127 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 16 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_meta-llama__Llama-2-7b-hf.0 likes2.3k downloads3y agoHugging Face08TTS-AGI /majestrino-unified-detailed-captions Majestrino Unified Detailed Captions Filtered subset of laion/majestrino-data containing all samples with unified_detailed_caption. Stats 4,658,407 samples 932 tar files (~1.1 GB each) ~1,017 GB total Format Each tar contains paired .flac + .json files. JSON fields: caption — the unified detailed caption caption_type — always unified_detailed_caption transcription — speech transcription (when available, normalized from multiple source keys) duration — audio… See the full description on the dataset page: https://huggingface.co/datasets/TTS-AGI/majestrino-unified-detailed-captions.audioaudio-classification1M<n<10M3 likes2.1k downloads6mo agoHugging Face09OALL /details_grimjim__Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge Dataset Card for Evaluation run of grimjim/Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge Dataset automatically created during the evaluation run of model grimjim/Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_grimjim__Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge.tabular100K<n<1M0 likes2.1k downloads2y agoHugging Face10open-llm-leaderboard-old /details_one-man-army__UNA-34Beagles-32K-bf16-v1 Dataset Card for Evaluation run of one-man-army/UNA-34Beagles-32K-bf16-v1 Dataset automatically created during the evaluation run of model one-man-army/UNA-34Beagles-32K-bf16-v1 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_one-man-army__UNA-34Beagles-32K-bf16-v1.0 likes1.8k downloads3y agoHugging Face11open-llm-leaderboard-old /details_EleutherAI__gpt-j-6b Dataset Card for Evaluation run of EleutherAI/gpt-j-6b Dataset Summary Dataset automatically created during the evaluation run of model EleutherAI/gpt-j-6b on the Open LLM Leaderboard. The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 8 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_EleutherAI__gpt-j-6b.0 likes1.6k downloads3y agoHugging Face12open-llm-leaderboard-old /details_declare-lab__starling-7B Dataset Card for Evaluation run of declare-lab/starling-7B Dataset automatically created during the evaluation run of model declare-lab/starling-7B on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_declare-lab__starling-7B.0 likes1.5k downloads3y agoHugging Face13OALL /details_SenseLLM__ReflectionCoder-DS-33B Dataset Card for Evaluation run of SenseLLM/ReflectionCoder-DS-33B Dataset automatically created during the evaluation run of model SenseLLM/ReflectionCoder-DS-33B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_SenseLLM__ReflectionCoder-DS-33B.tabular100K<n<1M0 likes1.2k downloads2y agoHugging Face14open-llm-leaderboard-old /details_princeton-nlp__Sheared-LLaMA-1.3B Dataset Card for Evaluation run of princeton-nlp/Sheared-LLaMA-1.3B Dataset Summary Dataset automatically created during the evaluation run of model princeton-nlp/Sheared-LLaMA-1.3B on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_princeton-nlp__Sheared-LLaMA-1.3B.0 likes1.2k downloads3y agoHugging Face15OALL /details_migtissera__Tess-M-v1.3 Dataset Card for Evaluation run of migtissera/Tess-M-v1.3 Dataset automatically created during the evaluation run of model migtissera/Tess-M-v1.3. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_migtissera__Tess-M-v1.3.tabular100K<n<1M0 likes1.2k downloads2y agoHugging Face16open-llm-leaderboard-old /details_PocketDoc__Dans-TotSirocco-7b Dataset Card for Evaluation run of PocketDoc/Dans-TotSirocco-7b Dataset Summary Dataset automatically created during the evaluation run of model PocketDoc/Dans-TotSirocco-7b on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_PocketDoc__Dans-TotSirocco-7b.0 likes1.2k downloads3y agoHugging Face17open-llm-leaderboard-old /details_togethercomputer__RedPajama-INCITE-7B-Base Dataset Card for Evaluation run of togethercomputer/RedPajama-INCITE-7B-Base Dataset Summary Dataset automatically created during the evaluation run of model togethercomputer/RedPajama-INCITE-7B-Base on the Open LLM Leaderboard. The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_togethercomputer__RedPajama-INCITE-7B-Base.0 likes1.1k downloads3y agoHugging Face18OALL /details_princeton-nlp__Llama-3-8B-ProLong-512k-Instruct Dataset Card for Evaluation run of princeton-nlp/Llama-3-8B-ProLong-512k-Instruct Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-8B-ProLong-512k-Instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_princeton-nlp__Llama-3-8B-ProLong-512k-Instruct.tabular100K<n<1M0 likes1.1k downloads2y agoHugging Face19OALL /details_Nexusflow__Athene-70B Dataset Card for Evaluation run of Nexusflow/Athene-70B Dataset automatically created during the evaluation run of model Nexusflow/Athene-70B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Nexusflow__Athene-70B.tabular100K<n<1M0 likes1.1k downloads2y agoHugging Face20OALL /details_vilm__Quyen-Pro-Max-v0.1 Dataset Card for Evaluation run of vilm/Quyen-Pro-Max-v0.1 Dataset automatically created during the evaluation run of model vilm/Quyen-Pro-Max-v0.1. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_vilm__Quyen-Pro-Max-v0.1.tabular100K<n<1M0 likes1.1k downloads2y agoHugging Face21ixelszy /detailedBG_Loraimage1K<n<10K0 likes1.1k downloads3y agoHugging Face22open-llm-leaderboard-old /details_meta-llama__Llama-2-70b-hf Dataset Card for Evaluation run of meta-llama/Llama-2-70b-hf Dataset Summary Dataset automatically created during the evaluation run of model meta-llama/Llama-2-70b-hf on the Open LLM Leaderboard. The dataset is composed of 124 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 11 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_meta-llama__Llama-2-70b-hf.0 likes1.1k downloads3y agoHugging Face23netsol /resume-score-details Resume and Job Description Matching Dataset Overview This dataset contains 1,031 samples of resumes and job descriptions (JDs) generated and assessed using GPT-4o. The primary goal of this dataset is to evaluate the alignment between resumes and job descriptions, aiding in the study of resume relevance, skill alignment, and job fit scoring based on predefined criteria. Dataset Composition The dataset includes resumes matched with job descriptions, with the… See the full description on the dataset page: https://huggingface.co/datasets/netsol/resume-score-details.text-classification1K<n<10K7 likes1k downloads2y agoHugging Face24open-llm-leaderboard-old /details_Intel__neural-chat-7b-v3-1 Dataset Card for Evaluation run of Intel/neural-chat-7b-v3-1 Dataset Summary Dataset automatically created during the evaluation run of model Intel/neural-chat-7b-v3-1 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Intel__neural-chat-7b-v3-1.0 likes960 downloads3y agoHugging Face25open-llm-leaderboard-old /details_one-man-army__una-neural-chat-v3-3-P2-OMA Dataset Card for Evaluation run of one-man-army/una-neural-chat-v3-3-P2-OMA Dataset automatically created during the evaluation run of model one-man-army/una-neural-chat-v3-3-P2-OMA on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_one-man-army__una-neural-chat-v3-3-P2-OMA.0 likes958 downloads3y agoHugging Face26OALL /details_Azure99__blossom-v5.1-34b Dataset Card for Evaluation run of Azure99/blossom-v5.1-34b Dataset automatically created during the evaluation run of model Azure99/blossom-v5.1-34b. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Azure99__blossom-v5.1-34b.tabular100K<n<1M0 likes927 downloads2y agoHugging Face27zyxhhnkh /DetailVerifyBench DetailVerifyBench Project Page | Paper | GitHub DetailVerifyBench is a rigorous benchmark designed for dense hallucination localization in long image captions. It comprises 1,000 high-quality images across five distinct domains: Chart, Movie, Nature, Poster, and UI. With an average caption length of over 200 words and dense, token-level annotations of multiple hallucination types, it stands as a challenging benchmark for evaluating the precise hallucination localization capabilities… See the full description on the dataset page: https://huggingface.co/datasets/zyxhhnkh/DetailVerifyBench.imageimage-text-to-text1K<n<10K4 likes914 downloads6mo agoHugging Face28open-llm-leaderboard-old /details_AA051610__FT Dataset Card for Evaluation run of AA051610/FT Dataset automatically created during the evaluation run of model AA051610/FT on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_AA051610__FT.0 likes878 downloads3y agoHugging Face29open-llm-leaderboard-old /details_meta-llama__Llama-2-13b-hf Dataset Card for Evaluation run of meta-llama/Llama-2-13b-hf Dataset Summary Dataset automatically created during the evaluation run of model meta-llama/Llama-2-13b-hf on the Open LLM Leaderboard. The dataset is composed of 123 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 8 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_meta-llama__Llama-2-13b-hf.1 likes861 downloads3y agoHugging Face30amztheory /details_Qwen__Qwen2-1.5B-Instruct Dataset Card for Evaluation run of Qwen/Qwen2-1.5B-Instruct Dataset automatically created during the evaluation run of model Qwen/Qwen2-1.5B-Instruct. The dataset is composed of 117 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/amztheory/details_Qwen__Qwen2-1.5B-Instruct.text100K<n<1M0 likes846 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.