CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01janck /bigscience-lama Dataset Card for LAMA: LAnguage Model Analysis - a dataset for probing and analyzing the factual and commonsense knowledge contained in pretrained language models. @inproceedings{petroni2020how, title={How Context Affects Language Models' Factual Predictions}, author={Fabio Petroni and Patrick Lewis and Aleksandra Piktus and Tim Rockt{"a}schel and Yuxiang Wu and Alexander H. Miller and Sebastian Riedel}, booktitle={Automated Knowledge Base Construction}, year={2020}… See the full description on the dataset page: https://huggingface.co/datasets/janck/bigscience-lama.texttext-retrieval10K<n<100K1 likes497 downloads4y agoHugging Face02G-reen /big_settext100K<n<1M0 likes390 downloads9mo agoHugging Face03bigstupidhats /lmsys-chat-entext100K<n<1M0 likes318 downloads2y agoHugging Face04nyuuzyou /bigslide Dataset Card for Bigslide.ru Presentations Dataset Summary This dataset contains metadata and original files for 50,872 presentations from the bigslide.ru platform, a presentation storage and viewing service for school students. The dataset includes information such as presentation titles, URLs, download URLs, and extracted text content where available. Languages The dataset is multilingual, with Russian being the primary language. Other languages present… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/bigslide.texttext-classification10K<n<100K0 likes303 downloads2y agoHugging Face05bigstupidhats /wildchat-entabular100K<n<1M0 likes162 downloads2y agoHugging Face06bigscience-catalogue-data-dev /lm_code_github-eval_subsettext10K<n<100K2 likes139 downloads5y agoHugging Face07bigscience /collaborative_catalogtextn<1K1 likes120 downloads4y agoHugging Face08bigstupidhats /UltraMedicaltext100K<n<1M0 likes99 downloads2y agoHugging Face09bigstupidhats /MetaMathQAtext100K<n<1M0 likes90 downloads2y agoHugging Face10bigstupidhats /financial-instruction-aq22text100K<n<1M0 likes77 downloads2y agoHugging Face11EliMasonTech /BigSynthPiano Dataset Card for PercePiano - Natural Language Evaluations 1. Dataset Summary This dataset builds upon the original PercePiano dataset, which was designed for the automatic evaluation of piano performances. The original PercePiano dataset contains 1,202 segments of classical piano performances annotated by music experts across 19 distinct perceptual features. While the original dataset provides highly structured, numeric, and multi-level perceptual metrics (such as… See the full description on the dataset page: https://huggingface.co/datasets/EliMasonTech/BigSynthPiano.text1K<n<10K1 likes59 downloads4mo agoHugging Face12open-llm-leaderboard /bigscience__bloom-7b1-detailsgated Dataset Card for Evaluation run of bigscience/bloom-7b1 Dataset automatically created during the evaluation run of model bigscience/bloom-7b1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bigscience__bloom-7b1-details.tabular10K<n<100K0 likes49 downloads2y agoHugging Face13Skorcht /bigsynthtext1K<n<10K0 likes44 downloads2y agoHugging Face14open-llm-leaderboard /bigscience__bloom-3b-detailsgated Dataset Card for Evaluation run of bigscience/bloom-3b Dataset automatically created during the evaluation run of model bigscience/bloom-3b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bigscience__bloom-3b-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face15open-llm-leaderboard /bigscience__bloom-1b7-detailsgated Dataset Card for Evaluation run of bigscience/bloom-1b7 Dataset automatically created during the evaluation run of model bigscience/bloom-1b7 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bigscience__bloom-1b7-details.tabular10K<n<100K0 likes26 downloads2y agoHugging Face16open-llm-leaderboard /abacusai__bigstral-12b-32k-detailsgated Dataset Card for Evaluation run of abacusai/bigstral-12b-32k Dataset automatically created during the evaluation run of model abacusai/bigstral-12b-32k The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/abacusai__bigstral-12b-32k-details.tabular10K<n<100K0 likes25 downloads2y agoHugging Face17bigscience /bloom-book-promptstextn<1K1 likes24 downloads4y agoHugging Face18open-llm-leaderboard /bigscience__bloom-1b1-detailsgated Dataset Card for Evaluation run of bigscience/bloom-1b1 Dataset automatically created during the evaluation run of model bigscience/bloom-1b1 The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bigscience__bloom-1b1-details.tabular10K<n<100K0 likes24 downloads2y agoHugging Face19bigsmoke05 /optimized-solidity-datasettextn<1K1 likes12 downloads2y agoHugging Face20bigstupidhats /prm800k-phase1text10K<n<100K0 likes10 downloads2y agoHugging Face21bigstupidhats /math-shepherd-preprocesstext1M<n<10M0 likes9 downloads2y agoHugging Face22bigsmoke05 /research_papers_datasettextn<1K1 likes7 downloads2y agoHugging Face23BigSaba /mesu-train-1.2text10K<n<100K0 likes2 downloads4mo agoHugging Face24G-reen /big_set_gemma_it_formattext100K<n<1M0 likes1 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.