datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Raiden-DeepSeek-R1Click here to support our open-source dataset and model releases!
Raiden-DeepSeek-R1 is a dataset containing creative-reasoning and analytic-reasoning responses, testing the limits of DeepSeek R1's reasoning skills!
This dataset contains:
63k 'creative_content' and 'analytical_reasoning' prompts from microsoft/orca-agentinstruct-1M-v1, with all responses generated by deepseek-ai/DeepSeek-R1.
Responses demonstrate the reasoning capabilities of DeepSeek's 685b parameter R1 reasoning model.… See the full description on the dataset page: https://huggingface.co/datasets/sequelbox/Raiden-DeepSeek-R1.raiden_garments_folding_baseline_a2
raiden_garments_folding_baseline_a2
LeRobot v2.1 dataset of bimanual YAM (Raiden) garment folding, converted with per-subtask language labels (cell A2).
129 episodes / 95,762 frames / 30 fps
21 unique task strings
Cameras: observation.images.top, left_wrist, right_wrist (224×224, AV1)
State/action: 14-D joints (left 6+gripper, right 6+gripper)
Each LeRobot episode is one labeled subtask (not the high-level collect prompt). Typical sequence per garment:
single out the… See the full description on the dataset page: https://huggingface.co/datasets/Sshawnin/raiden_garments_folding_baseline_a2.Genshin_Impact_RaidenShogun_Voice_koreanRaiden-DeepSeek-R1-PREVIEWThis is a preview of the full Raiden-Deepseek-R1 creative and analytical reasoning dataset, containing the first ~6k rows. Get the full dataset here!
This dataset uses synthetic data generated by deepseek-ai/DeepSeek-R1.
The initial release of Raiden uses 'creative_content' and 'analytical_reasoning' prompts from microsoft/orca-agentinstruct-1M-v1.
Dataset has not been reviewed for format or accuracy. All responses are synthetic and provided without editing.
Use as you will.
details_Kquant03__Raiden-16x3.43B
Dataset Card for Evaluation run of Kquant03/Raiden-16x3.43B
Dataset automatically created during the evaluation run of model Kquant03/Raiden-16x3.43B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Kquant03__Raiden-16x3.43B.BraTS-2024-Complete
BraTS 2024 Complete Prepared Dataset
Brain Tumor Segmentation (Leave a like 💖 if this helped you)
Dataset Description
This is an organized and verified version of the BraTS 2024 challenge datasets, including three tumor types.
Included Datasets
Dataset
Type
Cases
Source
BraTS-GLI
Glioma
1,809
Synapse (Dec 2024)
BraTS-MEN-RT
Meningioma + RT
571
Synapse (Feb 2025)
BraTS-PED
Pediatric
348
Cancer Imaging Archive… See the full description on the dataset page: https://huggingface.co/datasets/RaidenShogunUltimate/BraTS-2024-Complete.Raiden-Mini-DeepSeek-V3.2-SpecialeClick here to support our open-source dataset and model releases!
Raiden-Mini-DeepSeek-V3.2.Speciale is a dataset containing creative-reasoning and analytic-reasoning responses, testing the limits of DeepSeek-V3.2.Speciale's reasoning skills!
This dataset contains:
a default subset of ~8k 'creative_content' and 'analytical_reasoning' prompts from sequelbox/Raiden-DeepSeek-R1, with all responses generated by DeepSeek V3.2 Speciale.
provides an unfiltered look into the reasoning skills of… See the full description on the dataset page: https://huggingface.co/datasets/sequelbox/Raiden-Mini-DeepSeek-V3.2-Speciale.raiden_shogun_genshin
Dataset of raiden_shogun/雷電将軍/雷电将军 (Genshin Impact)
This is the dataset of raiden_shogun/雷電将軍/雷电将军 (Genshin Impact), containing 500 images and their tags.
The core tags of this character are long_hair, purple_hair, purple_eyes, breasts, mole, mole_under_eye, large_breasts, hair_ornament, braid, very_long_hair, braided_ponytail, hair_flower, which are pruned in this dataset.
Images are crawled from many sites (e.g. danbooru, pixiv, zerochan ...), the auto-crawling system is powered… See the full description on the dataset page: https://huggingface.co/datasets/CyberHarem/raiden_shogun_genshin.sequelbox_Raiden-DeepSeek-R1-Shuffled-ShareGPTimport json
from tqdm import tqdm
from datasets import load_dataset
import pandas as pd
# Example usage:
dataset = load_dataset("sequelbox/Raiden-DeepSeek-R1")["train"]
dataset = dataset.shuffle(seed=42)
output_file = "./sequelbox_Raiden-DeepSeek-R1-Shuffled-ShareGPT.parquet"
data = []
for item in tqdm(dataset):
if item["prompt"].strip() == "" or item["response"].strip() == "":
continue
data.append(
{
"conversations": [
{… See the full description on the dataset page: https://huggingface.co/datasets/PJMixers-Dev/sequelbox_Raiden-DeepSeek-R1-Shuffled-ShareGPT.Raiden-DeepSeek-R1-llama3.1-v1raiden-fichier-consolide-bornes-de-recharge-pour-vehicules-electriques
RAIDEN - fichier consolidé - bornes de recharge pour véhicules électriques
[!NOTE]
Ce jeu de données Hugging Face est vide. Cette carte sert seulement à référencer le jeu de données **RAIDEN - fichier consolidé - bornes de recharge pour véhicules électriques ** qui est disponible à l'adresse https://www.data.gouv.fr/datasets/63f38ce81be1f6867a668b51
Description
Fichier consolidé qui contient l'ensemble des pdc de la société RAIDEN
License
Licence… See the full description on the dataset page: https://huggingface.co/datasets/french-open-data/raiden-fichier-consolide-bornes-de-recharge-pour-vehicules-electriques.sequelbox_Raiden-DeepSeek-R1-PREVIEW-Shuffled-ShareGPTRaiden-DeepSeek-R1-modtokenRaiden-DeepSeek-R1-llama3.1nva-Raidenjob-intelligence-artifactsCIFAR100CorridorKey-WebDatasetCloudSEN12-hqraiden_five_garments_baseline_round1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 104,
"total_frames": 67515,
"total_tasks": 11,
"total_videos": 312,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:104"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_baseline_round1.raiden_five_garments_baseline_round2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 150,
"total_frames": 90516,
"total_tasks": 12,
"total_videos": 450,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:150"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_baseline_round2.raiden_five_garments_baseline_round3This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 305,
"total_frames": 92897,
"total_tasks": 11,
"total_videos": 915,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:305"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_baseline_round3.raiden_five_garments_baseline_round4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 222,
"total_frames": 50596,
"total_tasks": 11,
"total_videos": 666,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:222"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_baseline_round4.raiden_five_garments_baseline_round5This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 223,
"total_frames": 52924,
"total_tasks": 11,
"total_videos": 669,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:223"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_baseline_round5.raiden_five_garments_adverserial_round1_2_3This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 582,
"total_frames": 253788,
"total_tasks": 12,
"total_videos": 1746,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:582"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_adverserial_round1_2_3.raiden_five_garments_adverserial_round4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 204,
"total_frames": 63670,
"total_tasks": 11,
"total_videos": 612,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:204"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_adverserial_round4.raiden_five_garments_adverserial_round5This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam",
"total_episodes": 202,
"total_frames": 59780,
"total_tasks": 11,
"total_videos": 606,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:202"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/huzheyuan/raiden_five_garments_adverserial_round5.
