CoolFace
17 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Pensioner /shotplan ShotPlan Training Dataset Multi-shot video training data for ShotPlan: Cinematic Video Generation with Learnable Planning Token. 💻 Code: https://github.com/Pensioner-11/ShotPlan 🤖 Models: ShotPlan-Wan2.1-T2V-14B · ShotPlan-Wan2.2-T2V-A14B-HighNoise Contents Path Description data/train_meta_16fps.json 6,404 training samples (metadata + captions) data/videos/V*_16fps.mp4 549 source videos, re-encoded to 16 fps Each sample is an 80-frame (5 s @… See the full description on the dataset page: https://huggingface.co/datasets/Pensioner/shotplan.tabulartext-to-video1K<n<10K7 likes365 downloads2mo agoHugging Face02nyu-dice-lab /lm-eval-results-penfever-Llama-3-8B-tulu-human-v2-private Dataset Card for Evaluation run of penfever/Llama-3-8B-tulu-human-v2 Dataset automatically created during the evaluation run of model penfever/Llama-3-8B-tulu-human-v2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Llama-3-8B-tulu-human-v2-private.tabular100K<n<1M0 likes259 downloads2y agoHugging Face03nyu-dice-lab /lm-eval-results-penfever-Llama-3-8B-NuminaCoT-private Dataset Card for Evaluation run of penfever/Llama-3-8B-NuminaCoT Dataset automatically created during the evaluation run of model penfever/Llama-3-8B-NuminaCoT The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Llama-3-8B-NuminaCoT-private.tabular100K<n<1M0 likes133 downloads2y agoHugging Face04Clinical-Reasoning-Hub /pentabrid-reproducibility Pentabrid 27B: reproducibility package Everything required to recompute the results of a controlled evaluation of fine-tuning configurations for medical question answering. Openly available with no access restrictions. Contents Path Description per_item/medxpertqa_*.jsonl Per-item predictions for all six checkpoints on 2,450 MedXpertQA-Text items. Fields: id, gold, extracted_answer, correct, explicit_marker_present, n_markers, response_chars… See the full description on the dataset page: https://huggingface.co/datasets/Clinical-Reasoning-Hub/pentabrid-reproducibility.tabularquestion-answering10K<n<100K0 likes123 downloads6d agoHugging Face05nyu-dice-lab /lm-eval-results-penfever-Amber-7B-000-tulu-v2-private Dataset Card for Evaluation run of penfever/Amber-7B-000-tulu-v2 Dataset automatically created during the evaluation run of model penfever/Amber-7B-000-tulu-v2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Amber-7B-000-tulu-v2-private.tabular100K<n<1M0 likes63 downloads2y agoHugging Face06Penguin-Scrolls /PenguinScrolls PenguinScrolls: A User-Aligned Fine-Grained Benchmark for Long-Context Language Model Evaluation Introduction PenguinScrolls (企鹅卷轴) is a comprehensive benchmark designed to evaluate and enhance the long-text processing capabilities of large language models (LLMs). Current benchmarks for evaluating long-context language models often rely on synthetic tasks that fail to adequately reflect real user needs, leading to a weak correlation between benchmark scores and actual… See the full description on the dataset page: https://huggingface.co/datasets/Penguin-Scrolls/PenguinScrolls.tabular1K<n<10K6 likes53 downloads2y agoHugging Face07nyu-dice-lab /lm-eval-results-penfever-Mistral-7B-tulu-v2-private Dataset Card for Evaluation run of penfever/Mistral-7B-tulu-v2 Dataset automatically created during the evaluation run of model penfever/Mistral-7B-tulu-v2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Mistral-7B-tulu-v2-private.tabular100K<n<1M0 likes40 downloads2y agoHugging Face08pensieves /gsm8ktabular1K<n<10K0 likes30 downloads2y agoHugging Face09pengxiang /commtokentabularn<1K0 likes24 downloads4mo agoHugging Face10open-llm-leaderboard /HoangHa__Pensez-Llama3.1-8B-detailsgated Dataset Card for Evaluation run of HoangHa/Pensez-Llama3.1-8B Dataset automatically created during the evaluation run of model HoangHa/Pensez-Llama3.1-8B The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HoangHa__Pensez-Llama3.1-8B-details.tabular10K<n<100K0 likes20 downloads2y agoHugging Face11open-llm-leaderboard /DreadPoor__Minus_Penus-8B-Model_Stock-detailsgated Dataset Card for Evaluation run of DreadPoor/Minus_Penus-8B-Model_Stock Dataset automatically created during the evaluation run of model DreadPoor/Minus_Penus-8B-Model_Stock The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Minus_Penus-8B-Model_Stock-details.tabular10K<n<100K0 likes15 downloads2y agoHugging Face12ksophiena /kb-yurisprudensi-pencuriantabular1K<n<10K0 likes15 downloads3mo agoHugging Face13fullneime /pencilboxtabularn<1K0 likes11 downloads3y agoHugging Face14open-llm-leaderboard /DreadPoor__tests_pending-do_not_use_yet-detailsgated Dataset Card for Evaluation run of DreadPoor/tests_pending-do_not_use_yet Dataset automatically created during the evaluation run of model DreadPoor/tests_pending-do_not_use_yet The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__tests_pending-do_not_use_yet-details.tabular10K<n<100K0 likes11 downloads2y agoHugging Face15joe-harris14 /pen-into-box pen-into-box This dataset was generated using a phospho starter pack. This dataset contains a series of episodes recorded with a robot and multiple cameras. It can be directly used to train a policy using imitation learning. It's compatible with LeRobot and RLDS. tabularroboticsn<1K0 likes7 downloads1y agoHugging Face16NathanGavenski /Pendulum-v1tabular100K<n<1M0 likes2 downloads2y agoHugging Face17PennyML /7578578testpmontabular100K<n<1M0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.