CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lilywchen /lucky-initialization-atlas-100m-v2 Lucky initialization atlas v2 evidence Private live evidence archive for lilywchen/lucky-initialization-atlas-100m-v2. It contains hash-bound configs, provenance, scalar trajectories, step-zero diagnostics, and final per-sequence losses after those artifacts complete. It excludes credentials, caches, raw FineWeb-derived token arrays, optimizer states, and W&B binary logs. tabulartext-generationn<1K0 likes330 downloads27d agoHugging Face02lilonghao /MM-ContextASR-Bench MM-ContextASR Bench Metadata and evaluation splits for Multimodal Conversational Context for LLM-Based ASR: Data Construction, Training, and Benchmark. Dataset summary Config Examples Audio Context Primary metric mm_contextasr 1,250 (250 current utterances × 5 histories) 1,439 WAV files included Controlled user-assistant dialogue entity Recall kespeech 19,212 Source ID only Same-speaker speech and transcript CER, SER, entity Recall cv_yue 3,525… See the full description on the dataset page: https://huggingface.co/datasets/lilonghao/MM-ContextASR-Bench.audioautomatic-speech-recognition10K<n<100K1 likes226 downloads7d agoHugging Face03li-lab /First-do-NOHARM-v2 Japanese First Do NOHARM v2 This dataset is a human-reviewed Japanese adaptation of the First Do NOHARM v2 benchmark for evaluating the safety of large language models in clinical decision-making. Dataset Overview The dataset contains 330 Japanese clinical prompts, consisting of: 30 baseline cases 300 perturbation items (10 perturbations per baseline case) Each baseline case and its 10 perturbations share the same rubric and matching guidance. Perturbations… See the full description on the dataset page: https://huggingface.co/datasets/li-lab/First-do-NOHARM-v2.tabularn<1K0 likes103 downloads13d agoHugging Face04lgomezjurado-lila /qwen3-orthdion-sweeptabular100K<n<1M0 likes67 downloads2mo agoHugging Face05lilbillbiscuit /biocoder_publictabularn<1K4 likes59 downloads2y agoHugging Face06lildosa /tamil-nadu-government-health-facilities Tamil Nadu Government Health Facilities A cleaned and structured dataset of government healthcare facilities across Tamil Nadu, India. Dataset Description This dataset contains 2,398 healthcare facility records across Tamil Nadu, including government hospitals, Primary Health Centres (PHCs), Urban Primary Health Centres (UPHCs), maternity facilities, dental facilities, and other government healthcare facilities. Each record includes information such as: Facility… See the full description on the dataset page: https://huggingface.co/datasets/lildosa/tamil-nadu-government-health-facilities.tabular1K<n<10K1 likes57 downloads1mo agoHugging Face07open-llm-leaderboard /LilRg__ECE_Finetunning-detailsgated Dataset Card for Evaluation run of LilRg/ECE_Finetunning Dataset automatically created during the evaluation run of model LilRg/ECE_Finetunning The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__ECE_Finetunning-details.tabular10K<n<100K0 likes46 downloads2y agoHugging Face08open-llm-leaderboard /Lil-R__PRYMMAL-ECE-1B-SLERP-V1-detailsgated Dataset Card for Evaluation run of Lil-R/PRYMMAL-ECE-1B-SLERP-V1 Dataset automatically created during the evaluation run of model Lil-R/PRYMMAL-ECE-1B-SLERP-V1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__PRYMMAL-ECE-1B-SLERP-V1-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face09open-llm-leaderboard /Lil-R__2_PRYMMAL-ECE-7B-SLERP-detailsgated Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face10open-llm-leaderboard /LilRg__PRYMMAL-6B-slerp-detailsgated Dataset Card for Evaluation run of LilRg/PRYMMAL-6B-slerp Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-6B-slerp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-6B-slerp-details.tabular10K<n<100K0 likes41 downloads2y agoHugging Face11li-lab /JP-AlpaCare-MedInstruct-52k JP-AlpaCare-MedInstruct-52k This dataset is a Japanese-translated and aligned version of AlpaCare-MedInstruct-52k. The translation was performed automatically using gpt-4o-2024-05-13, preserving alignment between English and Japanese instructions, inputs, and outputs. Total data size is 51992. Dataset Details Original Dataset: AlpaCare-MedInstruct-52k Translation Model: GPT-4o (gpt-4o-2024-05-13) Fields: id (ID) instruction_ja, input_ja, output_ja (Japanese) id_en… See the full description on the dataset page: https://huggingface.co/datasets/li-lab/JP-AlpaCare-MedInstruct-52k.tabular10K<n<100K0 likes41 downloads1y agoHugging Face12open-llm-leaderboard /LilRg__PRYMMAL-ECE-7B-SLERP-V3-detailsgated Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V3 Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V3-details.tabular10K<n<100K0 likes34 downloads2y agoHugging Face13open-llm-leaderboard /LilRg__PRYMMAL-ECE-7B-SLERP-V6-detailsgated Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V6 Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V6 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V6-details.tabular10K<n<100K0 likes34 downloads2y agoHugging Face14open-llm-leaderboard /LilRg__PRYMMAL-ECE-7B-SLERP-V7-detailsgated Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V7 Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V7 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V7-details.tabular10K<n<100K0 likes33 downloads2y agoHugging Face15open-llm-leaderboard /Lil-R__2_PRYMMAL-ECE-7B-SLERP-V1-detailsgated Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP-V1 Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP-V1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-V1-details.tabular10K<n<100K0 likes33 downloads2y agoHugging Face16open-llm-leaderboard /LilRg__PRYMMAL-ECE-7B-SLERP-V5-detailsgated Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V5 Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V5 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V5-details.tabular10K<n<100K0 likes32 downloads2y agoHugging Face17open-llm-leaderboard /Lil-R__2_PRYMMAL-ECE-7B-SLERP-V2-detailsgated Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP-V2 Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP-V2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-V2-details.tabular10K<n<100K0 likes32 downloads2y agoHugging Face18open-llm-leaderboard /Lil-R__PRYMMAL-ECE-7B-SLERP-V8-detailsgated Dataset Card for Evaluation run of Lil-R/PRYMMAL-ECE-7B-SLERP-V8 Dataset automatically created during the evaluation run of model Lil-R/PRYMMAL-ECE-7B-SLERP-V8 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__PRYMMAL-ECE-7B-SLERP-V8-details.tabular10K<n<100K0 likes32 downloads2y agoHugging Face19open-llm-leaderboard /LilRg__PRYMMAL-ECE-7B-SLERP-V4-detailsgated Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V4 Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V4-details.tabular10K<n<100K0 likes31 downloads2y agoHugging Face20open-llm-leaderboard /Lil-R__2_PRYMMAL-ECE-7B-SLERP-V3-detailsgated Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP-V3 Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP-V3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-V3-details.tabular10K<n<100K0 likes31 downloads2y agoHugging Face21Lilambd /japan-dpc-hospitals Japan DPC Hospitals — all 4414 acute-care hospitals Every hospital in Japan's DPC (Diagnosis Procedure Combination) inpatient payment survey, in one clean table: name, prefecture, municipality code, beds and annual discharges. Source: the Ministry of Health, Labour and Welfare (MHLW) FY2024 (令和6年度) DPC discharge-patient survey, as published. Who uses this: medical-device and pharma sales teams sizing territories, market analysts ranking hospitals by volume, researchers who need… See the full description on the dataset page: https://huggingface.co/datasets/Lilambd/japan-dpc-hospitals.tabular1K<n<10K0 likes30 downloads2d agoHugging Face22Lilambd /japan-prefecture-facts Japan by Prefecture — 564 sourced facts for all 47 prefectures Regional minimum wage (FY2024), jobs-to-applicants ratio, consumer price regional difference index (overall and by category, 2024), foreign residents (2024), and a derived real minimum wage (minimum wage ÷ regional price index × 100). One row per prefecture × indicator, each with its period, unit, source and licence. Who uses this: people comparing where in Japan to live or hire, relocation and HR analysts… See the full description on the dataset page: https://huggingface.co/datasets/Lilambd/japan-prefecture-facts.textn<1K0 likes27 downloads2d agoHugging Face23Lilith88 /BeingAFriendlyYetHumorusChatBotTruthfully, this is less about the data and more about how I collected it. This is synthetic data and I am currently in the process of creating an all-in-one synthetic data manufacturing suite. Care to check it out for yourself? https://ai.studio/apps/25f0d17c-a5a3-4fee-884d-b5935b48f222?fullscreenApplet=true tabularn<1K0 likes17 downloads2mo agoHugging Face24open-llm-leaderboard /DoppelReflEx__MN-12B-LilithFrame-detailsgated Dataset Card for Evaluation run of DoppelReflEx/MN-12B-LilithFrame Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-LilithFrame The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-LilithFrame-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face25open-llm-leaderboard /DoppelReflEx__MN-12B-LilithFrame-Experiment-2-detailsgated Dataset Card for Evaluation run of DoppelReflEx/MN-12B-LilithFrame-Experiment-2 Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-LilithFrame-Experiment-2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-LilithFrame-Experiment-2-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face26open-llm-leaderboard /DoppelReflEx__MN-12B-LilithFrame-Experiment-3-detailsgated Dataset Card for Evaluation run of DoppelReflEx/MN-12B-LilithFrame-Experiment-3 Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-LilithFrame-Experiment-3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-LilithFrame-Experiment-3-details.tabular10K<n<100K0 likes10 downloads2y agoHugging Face27open-llm-leaderboard /DoppelReflEx__MN-12B-LilithFrame-Experiment-4-detailsgated Dataset Card for Evaluation run of DoppelReflEx/MN-12B-LilithFrame-Experiment-4 Dataset automatically created during the evaluation run of model DoppelReflEx/MN-12B-LilithFrame-Experiment-4 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DoppelReflEx__MN-12B-LilithFrame-Experiment-4-details.tabular10K<n<100K0 likes9 downloads2y agoHugging Face28build-small-hackathon /forgotten-lily-traces Forgotten Lily — Gameplay Traces Anonymous turn-by-turn traces from Forgotten Lily, a narrative mystery game built for the Hugging Face Build Small hackathon (Thousand Token Wood). In the game you play a detective questioning Lily, a girl who can no longer speak in words — only in tones, a private language of 28 glyphs that each carry one fragment of meaning. This dataset captures, for each turn: the question the player asked, the tones Lily answered with (their IDs and glyphs)… See the full description on the dataset page: https://huggingface.co/datasets/build-small-hackathon/forgotten-lily-traces.tabularn<1K0 likes8 downloads3mo agoHugging Face29open-llm-leaderboard /Lil-R__2_PRYMMAL-ECE-2B-SLERP-V2-detailsgated Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-2B-SLERP-V2 Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-2B-SLERP-V2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-2B-SLERP-V2-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face30open-llm-leaderboard /LilRg__10PRYMMAL-3B-slerp-detailsgated Dataset Card for Evaluation run of LilRg/10PRYMMAL-3B-slerp Dataset automatically created during the evaluation run of model LilRg/10PRYMMAL-3B-slerp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__10PRYMMAL-3B-slerp-details.tabular10K<n<100K0 likes6 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.