datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
v100-ds-mmluwarehouse_floor1_rgb_test_30fps_v100This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Jeonminjun/warehouse_floor1_rgb_test_30fps_v100.so101_test_task_v100This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 101,
"total_frames": 56215,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:101"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/minhducvm1912/so101_test_task_v100.eval_red_on_blue_ACT_v100This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 2,
"total_frames": 11,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:2"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Lakesenberg/eval_red_on_blue_ACT_v100.qwen-image-v100
Qwen-Image-2.1 on a Tesla V100 32 GB
Comparison images for a write-up on running Qwen-Image-2.1 with stable-diffusion.cpp on a Tesla V100 PCIe 32 GB.
Prompt for both: "a neon sign that reads QWEN on a rainy city street at night, wet pavement reflections, cinematic, highly detailed", 1024×1024, cfg 6.0, seed 42.
cache-vs-steps-6dcb5bb.jpg: 40 steps without caching, 40 steps with EasyCache, and 20 steps without caching (sd.cpp 6dcb5bb, Q4_K, euler).
samplers-2f88688.jpg: euler… See the full description on the dataset page: https://huggingface.co/datasets/elpatron2/qwen-image-v100.Sheetpedia_json_v1006continual-internalization
continual-internalization/benchmark
Aggregated benchmark across three continual-internalization settings:
world-news — Polymarket-spike-anchored news articles (Feb–Mar 2026), post-cutoff.
code-changelogs — new public Python APIs introduced in stable releases of NumPy / pandas / Polars / PyTorch / SciPy.
personalization — PersonaMem-v2 (static, K=1) + HorizonBench (streaming, K=4) persona conversations.
Splits
evaluation
Eval questions only. Schema:… See the full description on the dataset page: https://huggingface.co/datasets/anon-neurips-2026-v100/continual-internalization.simon-arc-combine-v100
Version 1
A combination of multiple datasets.
Datasets: dataset_solve_color.jsonl, dataset_solve_rotate.jsonl, dataset_solve_translate.jsonl.
Version 2
Datasets: dataset_solve_color.jsonl, dataset_solve_rotate.jsonl, dataset_solve_translate.jsonl.
Version 3
Datasets: dataset_solve_color.jsonl, dataset_solve_rotate.jsonl, dataset_solve_translate.jsonl.
Version 4
Added a shared dataset name for all these datasets: SIMON-SOLVE-V1. There may be higher… See the full description on the dataset page: https://huggingface.co/datasets/neoneye/simon-arc-combine-v100.ms_swift_0304_v100K_t200Kclean_dirty_dac_test_complex_v100clean_dirty_dac_complex_v100V1000
