datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DREAM-1K
DREAM-1K
DREAM-1K (Description
with Rich Events, Actions, and Motions) is a challenging video description benchmark. It contains a collection of 1,000 short (around 10 seconds) video clips with diverse complexities from five different origins: live-action movies, animated movies, stock videos, long YouTube videos, and TikTok-style short videos. We provide a fine-grained manual annotation for each video.
Bellow is the dataset statistics:
bill_summary_us
Dataset Card for "bill_summary_us"
Dataset Summary
Dataset for summarization of summarization of US Congressional bills (bill_summary_us).
Supported Tasks and Leaderboards
More Information Needed
Languages
English
Dataset Structure
Data Instances
default
Data Fields
id: id of the bill in format(congress number + bill type + bill number + bill version).
congress: number of the congress.
bill_type: type of… See the full description on the dataset page: https://huggingface.co/datasets/dreamproit/bill_summary_us.lm-eval-results-DreadPoor-Harpy-7B-Model_Stock-private
Dataset Card for Evaluation run of DreadPoor/Harpy-7B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Harpy-7B-Model_Stock
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-DreadPoor-Harpy-7B-Model_Stock-private.bill_text_us
Dataset Card for "bill_text_us"
Dataset Summary
Dataset for US Congressional bills (bill_text_us).
Supported Tasks and Leaderboards
More Information Needed
Languages
English
Dataset Structure
Data Instances
default
Data Fields
id: id of the bill in format(congress number + bill type + bill number + bill version).
congress: number of the congress.
bill_type: type of the bill.
bill_number: number of the… See the full description on the dataset page: https://huggingface.co/datasets/dreamproit/bill_text_us.dream-machine-ai
██████╗ █████╗ ████████╗ █████╗ ███████╗███████╗████████╗
██╔══██╗██╔══██╗╚══██╔══╝██╔══██╗██╔════╝██╔════╝╚══██╔══╝
██║ ██║███████║ ██║ ███████║███████╗█████╗ ██║
██║ ██║██╔══██║ ██║ ██╔══██║╚════██║██╔══╝ ██║
██████╔╝██║ ██║ ██║ ██║ ██║███████║███████╗ ██║
╚═════╝ ╚═╝ ╚═╝ ╚═╝ ╚═╝ ╚═╝╚══════╝╚══════╝ ╚═╝
✨ Dream-Machine · E-Commerce Synthetic Dataset ✨
1,000 Products · 1,000 Users · 5,000 Behaviors · JD /… See the full description on the dataset page: https://huggingface.co/datasets/dream-machine-ai/dream-machine-ai.dream-of-the-red-chamber-continuations
红楼梦续写 · Dream of the Red Chamber: 100 AI Continuations
项目简介
本数据集包含 92 个独立的AI续写版本,续写中国古典文学巅峰之作《红楼梦》的第八十一回至第一百零八回(共28回)。所有续写严格遵循曹雪芹前八十回中埋下的伏笔、谶语和人物命运,完全拒绝高鹗续书。
为什么做这个数据集
《红楼梦》的结局是世界文学史上最大的悬案之一。曹雪芹约于1763年去世前未能完成全书,仅留下前八十回。1791年左右,高鹗发表了一百二十回本,补写了后四十回,但红学研究日益表明高鹗续书严重违背了曹雪芹在前八十回中精心布置的伏笔。
曹雪芹原意 vs 高鹗续书
情节
曹雪芹原意
高鹗续书
黛玉之死
泪尽而亡,呼应"绛珠还泪"神话
焚稿断痴情
宝玉宝钗婚姻
"纵然是齐眉举案,到底意难平"
掉包计骗婚
贾府败落
政治牵连,锦衣军抄家,"忽喇喇似大厦倾"
败而复兴,"兰桂齐芳"
结局… See the full description on the dataset page: https://huggingface.co/datasets/PursuitOfDataScience/dream-of-the-red-chamber-continuations.trace-demomsmarco-2.1-segmentedbill_labels_us
Dataset Card for "bill_labels_us"
Dataset Summary
Dataset for US Congressional bills with policy area and legislative subjects information (bill_labels_us). Contains data for bills from the 108th to the 118th Congress, approximately 119,000 documents.
Supported Tasks and Leaderboards
More Information Needed
Languages
English
Dataset Structure
Data Instances
default
Data Fields
id: id of the bill in… See the full description on the dataset page: https://huggingface.co/datasets/dreamproit/bill_labels_us.DreadPoor__Promissum_Mane-8B-LINEAR-lorablated-details
Dataset Card for Evaluation run of DreadPoor/Promissum_Mane-8B-LINEAR-lorablated
Dataset automatically created during the evaluation run of model DreadPoor/Promissum_Mane-8B-LINEAR-lorablated
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Promissum_Mane-8B-LINEAR-lorablated-details.DreadPoor__BaeZel_V3-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/BaeZel_V3-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/BaeZel_V3-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__BaeZel_V3-8B-Model_Stock-details.DreadPoor__Heart_Stolen-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Heart_Stolen-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Heart_Stolen-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Heart_Stolen-8B-Model_Stock-details.DreadPoor__Summer_Rain-8B-TIES-details
Dataset Card for Evaluation run of DreadPoor/Summer_Rain-8B-TIES
Dataset automatically created during the evaluation run of model DreadPoor/Summer_Rain-8B-TIES
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Summer_Rain-8B-TIES-details.DreadPoor__WIP_Damascus-8B-TIES-details
Dataset Card for Evaluation run of DreadPoor/WIP_Damascus-8B-TIES
Dataset automatically created during the evaluation run of model DreadPoor/WIP_Damascus-8B-TIES
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__WIP_Damascus-8B-TIES-details.SGOCR
SGOCR
SGOCR is a spatially-grounded OCR visual question answering dataset for training and evaluating models that must read, localize, and reason about text in images.
The dataset contains grounded question-answer pairs over ChartQA, TextOCR, and COCO/COCO-Text source images. It is designed for OCR-aware VQA, text grounding, region-conditioned QA, and data-centric experiments around scene text understanding.
Project repository:… See the full description on the dataset page: https://huggingface.co/datasets/dreeseaw/SGOCR.DreadPoor__Aspire_V2-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Aspire_V2-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Aspire_V2-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Aspire_V2-8B-Model_Stock-details.DreamCubedHumanSample
Dream-Cubed Human Representative Sample
This directory is a deterministic, class-stratified sample of the Dream-Cubed Human dataset. It is provided for reviewer inspection of data quality; the full dataset remains available at https://huggingface.co/datasets/dream-cubed/DreamCubedHumanSample.
The sample dataset for procedurally generated data is available at https://huggingface.co/datasets/dream-cubed/DreamCubedNatural.
Depending on which commands have been run, the sample may… See the full description on the dataset page: https://huggingface.co/datasets/dream-cubed/DreamCubedHumanSample.DreadPoor__felix_dies-mistral-7B-model_stock-details
Dataset Card for Evaluation run of DreadPoor/felix_dies-mistral-7B-model_stock
Dataset automatically created during the evaluation run of model DreadPoor/felix_dies-mistral-7B-model_stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__felix_dies-mistral-7B-model_stock-details.DreadPoor__BaeZel-8B-LINEAR-details
Dataset Card for Evaluation run of DreadPoor/BaeZel-8B-LINEAR
Dataset automatically created during the evaluation run of model DreadPoor/BaeZel-8B-LINEAR
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__BaeZel-8B-LINEAR-details.DreadPoor__Zelus-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Zelus-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Zelus-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Zelus-8B-Model_Stock-details.eval_actuator_unboxing_pi05_sweep_v2_01_freeze_01_fullDreadPoor__Emu_Eggs-9B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Emu_Eggs-9B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Emu_Eggs-9B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Emu_Eggs-9B-Model_Stock-details.DreadPoor__Mercury_In_Retrograde-8b-Model-Stock-details
Dataset Card for Evaluation run of DreadPoor/Mercury_In_Retrograde-8b-Model-Stock
Dataset automatically created during the evaluation run of model DreadPoor/Mercury_In_Retrograde-8b-Model-Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Mercury_In_Retrograde-8b-Model-Stock-details.DreadPoor__TEST03-ignore-details
Dataset Card for Evaluation run of DreadPoor/TEST03-ignore
Dataset automatically created during the evaluation run of model DreadPoor/TEST03-ignore
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__TEST03-ignore-details.DreadPoor__Condensed_Milk-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Condensed_Milk-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Condensed_Milk-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Condensed_Milk-8B-Model_Stock-details.DreadPoor__BaeZel-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/BaeZel-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/BaeZel-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__BaeZel-8B-Model_Stock-details.DreadPoor__Nother_One-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Nother_One-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Nother_One-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Nother_One-8B-Model_Stock-details.DreadPoor__Rusted_Platinum-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/Rusted_Platinum-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/Rusted_Platinum-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__Rusted_Platinum-8B-Model_Stock-details.DreadPoor__ONeil-model_stock-8B-details
Dataset Card for Evaluation run of DreadPoor/ONeil-model_stock-8B
Dataset automatically created during the evaluation run of model DreadPoor/ONeil-model_stock-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__ONeil-model_stock-8B-details.DreadPoor__OrangeJ-8B-Model_Stock-details
Dataset Card for Evaluation run of DreadPoor/OrangeJ-8B-Model_Stock
Dataset automatically created during the evaluation run of model DreadPoor/OrangeJ-8B-Model_Stock
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/DreadPoor__OrangeJ-8B-Model_Stock-details.
