datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
liberoThis dataset was created using LeRobot.
Dataset Description
This dataset combines four individual Libero datasets: Libero-Spatial, Libero-Object, Libero-Goal and Libero-10.
All datasets were taken from here and converted into LeRobot format.
Homepage: https://libero-project.github.io
Paper: https://arxiv.org/abs/2306.03310
License: CC-BY 4.0
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "panda",
"total_episodes": 1693… See the full description on the dataset page: https://huggingface.co/datasets/physical-intelligence/libero.assetssyntheticDocQA_artificial_intelligence_test_beirBEIR version of vidore/syntheticDocQA_artificial_intelligence_test.
OmniEgo
D1 Headset Egocentric Whole-body Dataset
D1 is a headset multi-camera human motion dataset for humanoid intelligence, embodied AI, whole-body motion understanding, and imitation learning.
Overview
The D1 dataset is exported from the D1 headset multi-camera human motion capture system developed by Delta Intelligence. Each recorded episode contains synchronized multi-view video streams and whole-body skeleton and headset pose data.
The dataset supports research… See the full description on the dataset page: https://huggingface.co/datasets/Delta-Intelligence/OmniEgo.Hausa
Hausa Ajami OCR Dataset
Ce dataset contient des paires image/transcription de manuscrits haoussa en écriture ajami (écriture arabe adaptée au haoussa).
Contenu
Chaque ligne du fichier data/train/metadata.jsonl correspond à une ligne de texte ajami segmentée, avec :
file_name : nom du fichier image correspondant (image de la ligne, recadrée)
transcript : translittération en écriture latine de la ligne
source : identifiant du manuscrit d'origine (voir tableau… See the full description on the dataset page: https://huggingface.co/datasets/IntelligenceResearchLab/Hausa.syntheticDocQA_artificial_intelligence_test_beirBEIR version of vidore/syntheticDocQA_artificial_intelligence_test.
aloha_pen_uncap_diverseThis dataset was created using LeRobot.
Dataset Description
This dataset is a lerobot conversion of the aloha_pen_uncap_diverse subset of BiPlay.
BiPlay contains 9.7 hours of bimanual data collected with an aloha robot at the RAIL lab @ UC Berkeley, USA. It contains 7023 clips, 2000 language annotations and 326 unique scenes.
Paper: https://huggingface.co/papers/2410.10088 Code: https://github.com/sudeepdasari/dit-policy If you use the dataset please cite:… See the full description on the dataset page: https://huggingface.co/datasets/physical-intelligence/aloha_pen_uncap_diverse.KSAFE-MM
KSAFE-MM
📑 Paper |
🛠️ Technical Blog
📢 News
⚡️ 2026/06/11: Released on Hugging Face 🤗
📑 2026/05/29: arXiv preprint released
📕 2026/05/20: Technical blog article published
⚠️ CONTENT WARNING
This dataset contains potentially harmful and sensitive visual and textual content across the following 11 safety risk categories:
Risk Domain
Categories
Content Safety Risks
Hate and Unfairness, Violence, Sexual, Self-harm
Socio-economic Risks
Political and… See the full description on the dataset page: https://huggingface.co/datasets/K-intelligence/KSAFE-MM.syntheticDocQA_artificial_intelligence_test
Dataset Description
This dataset is part of a topic-specific retrieval benchmark spanning multiple domains, which evaluates retrieval in more realistic industrial applications.
It includes documents about the Artificial Intelligence.
Data Collection
Thanks to a crawler (see below), we collected 1,000 PDFs from the Internet with the query ('artificial intelligence'). From these documents, we randomly sampled 1000 pages.
We associated these with 100 questions and answers… See the full description on the dataset page: https://huggingface.co/datasets/vidore/syntheticDocQA_artificial_intelligence_test.docqa_artificial_intelligence_beirThis is a copy of https://huggingface.co/datasets/jinaai/docqa_artificial_intelligence reformatted into the BEIR format. For any further information like license, please refer to the original dataset.
Disclaimer
This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/docqa_artificial_intelligence_beir.objaverse.data.intelligence
Objaverse Data Intelligence: Scene-Level Structural Analysis and Decomposition Metadata
Large-scale scene-level geometry, structure, and material analysis of ~680K Objaverse assets, designed for ML-ready filtering, structural decomposition, and dataset curation.
🎥 Watch Reel
Quick Navigation
Overview
Key Features
Dataset Structure
Additional Files
Dataset Statistics
Classification & Detection Quality
Attribution
Citation
License
Acknowledgements… See the full description on the dataset page: https://huggingface.co/datasets/milimlee-synth3d/objaverse.data.intelligence.Emotion.IntelligenceComfyAgent-David
ComfyAgent-David
Final experiment image artifacts grouped by method.
comfyclaw/: final ComfyClaw runs from comfy_agent_experiments_output
comfygems/: final ComfyGEMS runs from cemfygems_comparison_output
baseline/: final baseline runs from comfy_agent_baseline_experiments_output
Included files are limited to image outputs under results/images/ and detailed/.
Smoke tests, caches, and earlier trial runs are excluded.
piper_uncap_penThis dataset was created using LeRobot.
About Task
Task Objective: Pick up the pen from the desk and uncap the pen.
Operational Objects: Marker Pen
Operation Duration: Each operation takes approximately 15 to 20 seconds.
Recording Frequency: 15 Hz.
Robot Type: 7-DOF dual-arm Agilex 2 Pipers Desktop Robot.
End Effector: Gripper.
Dual-Arm Operation: Yes.
Image Resolution: 640x480.
Camera Positions: High; Low; Left; Right.
Data Content: • Robot's current state. • Robot's… See the full description on the dataset page: https://huggingface.co/datasets/io-intelligence/piper_uncap_pen.solar-MFI-imagesfunctional-grasp-demos
Selected Grasp Demos
Contents
Each case folder contains:
grasp_data.npz — Top-1 ranked grasp (pre_grasp_dofs, grasp_target_dofs, reward, z_lift)
image_grasp.png — AI-generated grasp image (input to perception pipeline)
debug_retarget.png — WujiHand retarget visualization (if available)
Cases
Case
Object
Scale
Reward
z_lift
paper_coffee_cup_rot090__grasp_06
Paper Coffee Cup
1.0
0.742
0.196
paper_cup_8_oz_rot000__grasp_06
Paper Cup 8 Oz
1.0
0.696… See the full description on the dataset page: https://huggingface.co/datasets/Genesis-Intelligence/functional-grasp-demos.physical-intelligencevArabic-Math-SFT
Arabic Math SFT
A Multimodal Arabic Mathematics Dataset for Supervised Fine-Tuning
Empowering Arabic AI with visual mathematical reasoning
Overview
Arabic-Math-SFT is a curated multimodal dataset designed for supervised fine-tuning of vision-language models on mathematical problem-solving in Arabic. Each sample pairs a geometric or algebraic diagram with an Arabic-language problem statement and its corresponding solution.
This dataset bridges a… See the full description on the dataset page: https://huggingface.co/datasets/Omartificial-Intelligence-Space/Arabic-Math-SFT.syntheticDocQA_artificial_intelligence_test_tesseractds_benchmark_upv_climactquran_warsh_maghribiPearl-vdr-ar-train-preprocessed
Pearl-vdr-ar-train-preprocessed
Arabic culturally-aligned, VDR-style (query, image, hard-negatives) triplets for training multimodal embedding models with Sentence Transformers.
Dataset structure
Each row contains:
Column
Type
Description
query
string
Arabic text question about the image
category
string
High-level Arab-culture topic (Music, Landmarks, Cuisine, ...)
country
string
Country the sample is anchored to (Algeria, Saudi Arabia, ...)
image
image… See the full description on the dataset page: https://huggingface.co/datasets/Omartificial-Intelligence-Space/Pearl-vdr-ar-train-preprocessed.sunday-intelligence-imojis-preprocessedsyntheticDocQA_artificial_intelligence_test_ocr_chunksyntheticDocQA_artificial_intelligence_test_captioningdocqa_artificial_intelligence
Creation
This dataset is build upon the corresponding dataset from the ViDoRe Benchmark. For more information regarding the filtering please read our paper or this discussion on github.
Disclaimer
This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/docqa_artificial_intelligence.physical-intelligenceThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "aloha",
"total_episodes": 26,
"total_frames": 18507,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 50,
"splits": {
"train": "0:26"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ying01/physical-intelligence.Pearl-vdr-ar-train-hard-mined
Pearl-vdr-ar-train-hard-mined
Arabic culturally-aligned Visual Document Retrieval (VDR) triplets with model-mined hard negatives, derived from Omartificial-Intelligence-Space/Pearl-vdr-ar-train-preprocessed by replacing its metadata-based negatives with the top-4 most similar non-matching images according to Omartificial-Intelligence-Space/Qwen3-VL-Embedding-2B-Arabic-VDR (v2, 3-epoch finetune).
Dataset structure
Each row contains:
Column
Type
Description… See the full description on the dataset page: https://huggingface.co/datasets/Omartificial-Intelligence-Space/Pearl-vdr-ar-train-hard-mined.wordpress-booking-appointment-plugins-market-intelligence-sample
WordPress Booking & Appointment Plugins Market Intelligence Dataset -- Free Evaluation Sample
This dataset packages public WordPress.org plugin-directory records for booking, appointment, and reservation plugins into one analysis-ready market-intelligence table.
Each row represents one WordPress plugin enriched with install and rating metrics, update recency, support-resolution signals, commercial-language flags, booking-workflow feature flags, buyer-segment heuristics, and… See the full description on the dataset page: https://huggingface.co/datasets/Karmane/wordpress-booking-appointment-plugins-market-intelligence-sample.syntheticDocQA_artificial_intelligence_test
Dataset Description
This dataset is part of a topic-specific retrieval benchmark spanning multiple domains, which evaluates retrieval in more realistic industrial applications.
It includes documents about the Artificial Intelligence.
Data Collection
Thanks to a crawler (see below), we collected 1,000 PDFs from the Internet with the query ('artificial intelligence'). From these documents, we randomly sampled 1000 pages.
We associated these with 100 questions and answers… See the full description on the dataset page: https://huggingface.co/datasets/Madhu348/syntheticDocQA_artificial_intelligence_test.
