flex
Datasets
All datasets matching “flex”robotwin_3d
RoboTwin 2.0 — 3D (RGB + Depth)
Bimanual manipulation data from the RoboTwin 2.0 simulator, in LeRobot v2.1 format, with per-camera ground-truth depth alongside RGB.
Tasks
50
Episodes
27,500 (550 per task, contiguous)
Frames
6,183,813
Robot
ALOHA-style bimanual, 14-DoF
Control rate
50 Hz
Cameras
3 (cam_high, cam_left_wrist, cam_right_wrist)
Resolution
240 × 320
Language instructions
1,039,891 unique corpus-wide; 100 entries per episode
Total size… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/robotwin_3d.robotwin_3d_text_embeds_cache
RoboTwin 2.0 3D — T5 Text Embedding Cache
Precomputed UMT5-XXL text embeddings for the
1,039,891 unique task prompts of the RoboTwin 2.0 3D dataset
(flex-pi/robotwin_3d), as consumed by
Wan2.2-TI2V-5B. Precomputing these costs substantial GPU time; this cache skips it.
Cache key
Task strings from meta/tasks.jsonl are wrapped in a fixed template before encoding, and the cache
key is the sha256 of that templated prompt:
DEFAULT_PROMPT = "A video recorded from a… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/robotwin_3d_text_embeds_cache.flexray-data
FleXray data release
Website: FleXray project page
Paper: FleXray: Universal Clinical X-ray Segmentation
Code: github.com/VictorButoi/FleXray
Training and evaluation data for FleXray, a pan-anatomy X-ray segmentation model.
This repository holds every real X-ray source whose license permits redistribution,
repackaged as image/mask pairs with fxr-dataset manifests and dataset-native
label names that FleXray maps into its common protocol (CC BY 4.0), plus two more parts of the… See the full description on the dataset page: https://huggingface.co/datasets/VictorButoi/flexray-data.FlexiSLM-Data-2M-s2s-compact
FlexiSLM-Data — Speech-to-Speech Part (2.43M filtered samples, 385G in size)
Paper: https://arxiv.org/abs/2606.31247
Demo page: https://flexislm.github.io/
Code: https://github.com/AmphionTeam/FlexiSLM
FlexiSLM-Data is a large-scale, single-turn English speech-to-speech dialogue dataset
for training FlexiSLM, a spoken language model.
This repository contains the paired prompt-and-response audio portion of the release in
WebDataset format.
Related data releases… See the full description on the dataset page: https://huggingface.co/datasets/FlexiSLM/FlexiSLM-Data-2M-s2s-compact.ljspeechThis is a public domain speech dataset consisting of 13,100 short audio
clips of a single speaker reading passages from 7 non-fiction books. A
transcription is provided for each clip. Clips vary in length from 1 to 10
seconds and have a total length of approximately 24 hours.libero_mujoco3.3.2_depth
LIBERO with depth (MuJoCo 3.3.2 re-render)
The four standard LIBERO benchmark suites,
re-rendered under MuJoCo 3.3.2 and republished in LeRobot v2.1 format with
per-camera metric depth added alongside the usual RGB. no_noops means idle
frames have been stripped.
Contents
This repo holds four independent LeRobot datasets, one per suite:
suite
episodes
frames
tasks
libero_10_no_noops_lerobot
388
104,160
10
libero_goal_no_noops_lerobot
433
52,895
10… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/libero_mujoco3.3.2_depth.
