datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DreamZero-DROID-DatadreamDREAM is a multiple-choice Dialogue-based REAding comprehension exaMination dataset. In contrast to existing reading comprehension datasets, DREAM is the first to focus on in-depth multi-turn multi-party dialogue understanding.PhysicalAI-Autonomous-Vehicle-Cosmos-Drive-Dreams
PhysicalAI-Autonomous-Vehicle-Cosmos-Drive-Dreams
Paper | Paper Website | GitHub
Download
We provide a download script to download our dataset. If you have enough space, you can use git to download a dataset from huggingface.
usage: download.py [-h] --odir ODIR
[--file_types {hdmap,lidar,synthetic}[,…]]
[--workers N] [--clean_cache]
required arguments:
--odir ODIR Output… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/PhysicalAI-Autonomous-Vehicle-Cosmos-Drive-Dreams.midi-classical-music
MIDI Classical Music
This dataset contains a comprehensive collection of MIDI files representing classical music compositions from various renowned composers.
The collection includes works from composers such as Bach, Beethoven, Chopin, Mozart, and many others.
The dataset is organized into directories by composer, with each directory containing MIDI files of their compositions.
The dataset is ideal for music analysis, machine learning models for music generation, and other… See the full description on the dataset page: https://huggingface.co/datasets/drengskapur/midi-classical-music.HR-Bench
Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models
🌐Homepage | 📖 Paper
📊 HR-Bench
We find that the highest resolution in existing multimodal benchmarks is only 2K. To address the current lack of high-resolution multimodal benchmarks, we construct HR-Bench. HR-Bench consists two sub-tasks: Fine-grained Single-instance Perception (FSP) and Fine-grained Cross-instance Perception (FCP).… See the full description on the dataset page: https://huggingface.co/datasets/DreamMr/HR-Bench.finqadataset_info:
features:
name: id
dtype: string
name: post_text
sequence: string
name: pre_text
sequence: string
name: question
dtype: string
name: answers
dtype: string
name: table
sequence:
sequence: string
splits:
name: train
num_bytes: 26984130
num_examples: 6251
name: validation
num_bytes: 3757103
num_examples: 883
name: test
num_bytes: 4838430
num_examples: 1147
download_size: 21240722
dataset_size: 35579663
vam-cross-evaluation-artifactsdreamzero-egoverse-360-pretrainmultispider
MultiSpider: Towards Benchmarking Multilingual Text-to-SQL Semantic Parsing
In this work, we present MultiSpider, a multilingual text-to-SQL dataset which covers seven languages (English, German, French, Spanish, Japanese, Chinese, and Vietnamese).
Find more details on paper and code.
Please be aware that the MultiSpider dataset is available in two versions: with_English_value and with_original_value. Our reported results are based on the with_English_value version to circumvent any… See the full description on the dataset page: https://huggingface.co/datasets/dreamerdeo/multispider.eclthdDreamCubedNatural
Dream-Cubed Natural
Dream-Cubed Natural is a dataset of procedurally generated Minecraft Java Edition v1.12.2 terrain represented as 32x32x32 voxel chunks. It is intended for research on controllable 3D generation, biome-conditioned voxel modeling, inpainting, outpainting, and synthetic environment generation.
This repository contains only the natural/procedural portion of Dream-Cubed. It includes raw chunks extracted from Minecraft worlds as well as processed class-conditional… See the full description on the dataset page: https://huggingface.co/datasets/dream-cubed/DreamCubedNatural.uscode
United States Code, versioned by release point
Every section of the United States Code, as published by the Office of the Law
Revision Counsel (OLRC) at uscode.house.gov, across
every release point from 113-21 (July 18, 2013) through the present. A release
point is OLRC's republication of the Code after a batch of Public Laws is
classified; this dataset covers 381 of them over 58 titles.
Each row carries the section's plain text, its verbatim USLM XML, its
citation, its place in… See the full description on the dataset page: https://huggingface.co/datasets/dreamproit/uscode.DreamTac
DreamTac → FiftyOne (Native Multimodal MCAP)
DreamTac from Peking
University, converted to native multimodal MCAP episodes.
A Franka Emika Panda works through contact-rich tabletop tasks while four
cameras record on one 20 fps clock: a third-person view, a wrist view, and two
Xense Photon vision-based tactile sensors mounted on the gripper fingertips.
The fingertips are the point of the release. Each is a gel pad printed with a
marker grid, and the grid deforms where the object… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/DreamTac.MusicNetDreamEditBench
DreamEditBench for Subject Replacement task and Subject Addition task.
The goal of subject replacement is to replace a subject from a source image with a customized subject. In contrast, the aim of the subject addition task is to add a customized
subject to a desired position in the source image. To standardize the evaluation of the two proposed tasks, we curate a new benchmark, i.e. DreamEditBench, consisting of 22 subjects in alignment with DreamBooth with 20 images for each… See the full description on the dataset page: https://huggingface.co/datasets/tianleliphoebe/DreamEditBench.Slakh2100-FLAC-Redux-Reduceddresscode_agnostic_and_densepose
DressCode Agnostic & DensePose Dataset
Agnostic images, corresponding masks, and DensePose images for the DressCode dataset.
Information about the usage can be found at:
https://github.com/jiwoohong93/ita-mdt_code
License
The Dress Code Dataset is proprietary to and © Yoox Net-a-Porter Group S.p.A. and its licensors.It is distributed by the University of Modena and Reggio Emilia and is available for non-commercial academic use under the licence terms provided… See the full description on the dataset page: https://huggingface.co/datasets/jiwoohong93/dresscode_agnostic_and_densepose.Liminal-Dreamcore-1K
Dreamcore
A Collection of 1000 AI-Generated Dreamcore Aesthetic Images
All images in this collection are AI-generated.
Architecture
The generation pipeline behind this collection:
What Is Dreamcore?
Dreamcore is an internet aesthetic that captures the visual language of dreams - specifically the strange, liminal, half-remembered quality of dream imagery. It sits in the same family as weirdcore, traumacore, and oddcore… See the full description on the dataset page: https://huggingface.co/datasets/luka0x12/Liminal-Dreamcore-1K.DreamCubedHuman
Dream-Cubed Human
Dream-Cubed Human is a dataset of voxelized Minecraft structures and terrain represented as 32x32x32 block chunks. It is intended for research on controllable 3D generation, Minecraft-like structure generation, inpainting, outpainting, and synthetic environment generation.
This repository contains raw chunks extracted from six human-authored Minecraft map sources and a processed augmented dataset used for training. The processed human_augmented_dataset/ also… See the full description on the dataset page: https://huggingface.co/datasets/dream-cubed/DreamCubedHuman.omni-dreams-scenes
AlpaDreams Sample Scenes
Sample scene dataset for the AlpaDreams autonomous vehicle world model.
Use of this dataset is governed by the NVIDIA Autonomous Vehicle Dataset License Agreement.
AgiBotWorld-Beta_G1_task_480_Making_sandwiches_with_salad_dressing_1213_version
agibot_task_480
This dataset converts the AgiBot format uniformly into LeRobot V3.0.
Dataset Statistics
robot_name: G1
end_effector: 夹爪
task: 用沙拉酱做三明治1213版
total_episodes: 916
total_tasks: 1
size: 84G
Dataset Structure
├── data
│ └── chunk-xxx
│ ├── file-xxx.parquet
├── meta
│ ├── episodes
│ │ └── chunk-xxx
│ │ └── file-xxx.parquet
│ ├── info.json
│ ├── stats.json
│ └── tasks.parquet
└── videos
├──… See the full description on the dataset page: https://huggingface.co/datasets/BAAI-DataCube/AgiBotWorld-Beta_G1_task_480_Making_sandwiches_with_salad_dressing_1213_version.omni-dreams-samples
AlpaDreams Samples
Curated single-view driving sequences for evaluating the
nvidia/alpadreams-dit world model.
Layout
data/
└── single_view/
├── <clip-id>/
| ├── <clip-id_...>.mp4 # ground truth video
│ ├── <clip-id_..._hdmap>.mp4 # HD-map rasterized conditioning video
│ ├── first_frame.png # RGB first frame, extracted from ground truth video
│ └── prompt.txt # text prompt
└──… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/omni-dreams-samples.dreambooth
Dataset Card for "dreambooth"
Dataset of the Google paper DreamBooth: Fine Tuning Text-to-Image Diffusion Models for Subject-Driven Generation
The dataset includes 30 subjects of 15 different classes. 9 out of these subjects are live subjects (dogs and cats) and 21 are objects. The dataset contains a variable number of images per subject (4-6). Images of the subjects are usually captured in different conditions, environments and under different angles.
We include a file… See the full description on the dataset page: https://huggingface.co/datasets/google/dreambooth.DreamOmni2Bench
DreamOmni2: Multimodal Instruction-based Editing and Generation Benchmark
This repository contains the DreamOmni2Bench benchmark dataset, introduced in the paper DreamOmni2: Multimodal Instruction-based Editing and Generation.
The DreamOmni2 project proposes two novel tasks: multimodal instruction-based editing and generation. These tasks support both text and image instructions and extend the scope to include both concrete and abstract concepts, greatly enhancing their practical… See the full description on the dataset page: https://huggingface.co/datasets/xiabs/DreamOmni2Bench.Dreamer-V1-DataAfter heavier cleaning, the remaining data size is 3.12M.
WebDreamer: Model-Based Planning for Web Agents
WebDreamer is a planning framework that enables efficient and effective planning for real-world web agent tasks. Check our paper for more details.
This work is a collaboration between OSUNLP and Orby AI.
Repository: https://github.com/OSU-NLP-Group/WebDreamer
Paper: https://arxiv.org/abs/2411.06559
Point of Contact: Kai Zhang
Models
Dreamer-7B:
General… See the full description on the dataset page: https://huggingface.co/datasets/osunlp/Dreamer-V1-Data.prof_report__Lykon-DreamShaper__multi__24
Dataset Card for "prof_report__Lykon-DreamShaper__multi__24"
More Information needed
trajectory_data_dream_32
d3LLM Trajectory Dataset
Project Page | Paper | GitHub | Blog
This repository contains the pseudo-trajectory distillation data used for training d3LLM (pseuDo-Distilled Diffusion Large Language Model), as introduced in the paper "d3LLM: Ultra-Fast Diffusion LLM using Pseudo-Trajectory Distillation".
Introduction
d3LLM is a framework designed to strike a balance between accuracy and parallelism in diffusion-based large language models (dLLMs). This dataset consists of… See the full description on the dataset page: https://huggingface.co/datasets/d3LLM/trajectory_data_dream_32.maniskill-dreamer4-expert
