datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
meta13sphere_IRS_DCE_Topological_Dynamics__Boundary_Dissolution_Physics
Resonance Resonance / IRS-DCE
MASTER README (FULL EXTENDED VERSION)
If you need the other data or pdf check on [https://huggingface.co/datasets/meta13sphere/phaseShift_shell_result_pdf]
[2026-09-22 Update]
MASK_BDP BBRCM: 조건부 운용의 통합 구조
배경·경계·분기·관측·계량을 분해하고 다시 조립하면서, 무엇이 보존되고 어떤 조건에서 결과가 달라지는지 정리한 연구입니다. 상보성 쌍과 제타/RH 관련 변환뿐 아니라 BBRCM 자체에도 같은 변형·스트레스 테스트를 적용했습니다.
주요 성과는 다음과 같습니다.
R32: 원래 함수공간과 정확한 직교사영 조건 아래에서 첫 셀의 잔차를 전체 극한 잔차와… See the full description on the dataset page: https://huggingface.co/datasets/meta13sphere/meta13sphere_IRS_DCE_Topological_Dynamics__Boundary_Dissolution_Physics.ritvij-saxena-iris-detection-pythonHere is the IRIS dataset for the project iris-detection-python.
Official Statement
I hereby declare that I do not own the rights to the dataset used in this project. This dataset was provided by the faculty and utilized solely for educational purposes as part of an assignment for the Biometrics course (CS 559) at the Illinois Institute of Technology.
The dataset is provided for academic and research purposes only, and I encourage others to use it responsibly for similar educational… See the full description on the dataset page: https://huggingface.co/datasets/saxenaritvij/ritvij-saxena-iris-detection-python.nih-chest-xray-14-flatVCog-Bench
What is the Visual Cognition Gap between Humans and Multimodal LLMs?
Description:
VCog-Bench is a publicly available zero-shot abstract visual reasoning (AVR) benchmark designed to evaluate Multimodal Large Language Models (MLLMs). This benchmark integrates two well-known AVR datasets from the AI community and includes a newly proposed MaRs-VQA dataset. The findings in VCog-Bench show that current state-of-the-art MLLMs and Vision-Language Models (VLMs), such as GPT-4o… See the full description on the dataset page: https://huggingface.co/datasets/IrohXu/VCog-Bench.RPX
RPX: Robot Perception X
RPX is a real-world RGB-D benchmark for measuring robot perception across scene changes. The canonical naren/all release combines the multi-object, egocentric, single-object, VQA, and tracking metadata that previously lived on separate dataset branches.
Code and benchmark toolkit: github.com/IRVLUTD/RPX
Recommended dataset revision: naren/all (pin the commit SHA printed by your download for reproducible results)
License: Creative Commons Attribution 4.0… See the full description on the dataset page: https://huggingface.co/datasets/IRVLUTD/RPX.malfunction-image-datasetIRIS-CloudDeep
IRIS-CloudDeep
Ground-based long-wave infrared (LWIR) images of the night sky, with the binary ground-truth masks and clear/cloud labels behind Sommer, Kabalan and Brunet (2025), Atmos. Meas. Tech. 18, 2083–2101.
An uncooled FLIR Tau2 microbolometer (640×512, 17 μm pitch, 8–14 μm band, 9 Hz) recorded two night-time campaigns in early 2023 at Prades-le-Lez, France (43°41′51″ N, 3°51′53″ E). A 60 mm f/1.25 lens gives a narrow imaging area of 10.4° × 8.3°, about 58″ per pixel. The… See the full description on the dataset page: https://huggingface.co/datasets/ASKabalan/IRIS-CloudDeep.IRIS
IRIS Dataset: Industrial Real-Sim Imagery Set
Overview
The IRIS Dataset is a comprehensive real-world dataset designed to study sim-to-real transfer for object detection in industrial robotic environments. This repository provides:
The complete real IRIS dataset: 508 annotated images of 32 mechanical components captured across four distinct, challenging industrial scenes.
Assets for synthetic data generation: All necessary 3D models, backgrounds, and materials to… See the full description on the dataset page: https://huggingface.co/datasets/Carraskito/IRIS.Iris_Database
Synthetic Iris Image Dataset
Overview
This repository contains a dataset of synthetic colored iris images generated using diffusion models based on our paper "Synthetic Iris Image Generation Using Diffusion Networks." The dataset comprises 17,695 high-quality synthetic iris images designed to be biometrically unique from the training data while maintaining realistic iris pigmentation distributions. In this repository we contain about 10000 filtered iris images with the… See the full description on the dataset page: https://huggingface.co/datasets/fatdove/Iris_Database.GenManip-Assets-IROS_AlohaRipVIS
RipVIS v1.8.4
This Readme describes the RipVIS dataset, its contents, structure, known limitations, how to use it and what to expect in future updates. For more details, future challenges and other information, keep an eye on RipVIS website or write to andrei.dumitriu@uni-wuerzburg.de .
Short description
RipVIS dataset was introduced with RipVIS: Rip Currents Video Instance Segmentation Benchmark for Beach Monitoring and Safety paper, accepted at CVPR 2025. It is the… See the full description on the dataset page: https://huggingface.co/datasets/Irikos/RipVIS.SynWBM
The SynWBM (Synthetic White Button Mushrooms) Dataset!
Synthetic dataset of white button mushrooms (Agaricus bisporus) with instance segmentation masks and depth maps.
Dataset Summary
The SynWBM Dataset is a collection of synthetic images of white button mushroom. The dataset incorporates rendered (using Blender) and generated (using Stable Diffusion XL) synthetic images for training mushroom segmentation models. Each image is annotated with instance segmentation masks… See the full description on the dataset page: https://huggingface.co/datasets/ABC-iRobotics/SynWBM.wildlife_in_irrigation_ponds
Wildlife in Irrigation Ponds Dataset
Dataset Summary
This dataset supports the training and evaluation of object detection models for monitoring irrigation ponds, with the goal of detecting people and animals that have fallen into the water. It comprises synthetically generated images produced using state-of-the-art diffusion models (Z-Image, FLUX), with a real photograph of a target irrigation pond used as the background.
The dataset includes four object classes… See the full description on the dataset page: https://huggingface.co/datasets/grupo-avispa/wildlife_in_irrigation_ponds.taco_play_testcodex-pets-sprite-sheets
Boogu Pets Sprite Sheets
Private dataset of 3,015 RGBA WebP sprite sheets used for the Boogu image-edit training experiments.
These sprite sheets were used to train the
Codex Pets Sprite Sheet Generator.
Example
The v1_max.webp sprite sheet:
Contents
images/: 3,015 original source sprite sheets in RGBA format. V1 sheets are 1536x1872 and V2 sheets are 1536x2288.
data/train/ and data/validation/: viewer-compatible copies of the same sheets, split… See the full description on the dataset page: https://huggingface.co/datasets/irotem98/codex-pets-sprite-sheets.ffhq_with_llava_shorter_captions
Dataset Card for "ffhq_with_llava_shorter_captions"
More Information needed
irsl_datadecide_resultsIRFL
Dataset Card for IRFL
Dataset Description
Leaderboards
Colab notebook code for IRFL evaluation
Languages
Dataset Structure
Data Fields
Dataset Creation
Considerations for Using the Data
Licensing Information
Citation Information
Dataset Description
The IRFL dataset consists of idioms, similes, metaphors with matching figurative and literal images, and two novel tasks of multimodal figurative detection and retrieval.Using human annotation and an automatic pipeline… See the full description on the dataset page: https://huggingface.co/datasets/lampent/IRFL.ReCamDriving-STORM-ExampleFRACTAL-IRGBThis repo contains the orthoimageries used to colorize point clouds in the FRACTAL semantic segmentation dataset.
They are provided as is, in their original size (50 x 50 m), spatial resolution (0.2 m, i.e. 250 x 250 pixels) and band order (near-infrared, red, green, blue).
They might be used for quick visual inspections of the FRACTAL's point clouds, or for more advanced use such as multimodal (2D-3D) deep learning training.
Note:
The files are ordered by filename. Inspecting a single zip… See the full description on the dataset page: https://huggingface.co/datasets/IGNF/FRACTAL-IRGB.TPSoSe2026_Dataset_Collection_LeRobot_SO101
SO-101 Early Collection (superseded)
The project's first, miscellaneous recordings, made before the team settled on four
fixed tasks, a consistent recording protocol, and systematic prompt variation.
[!WARNING]
This dataset is superseded and not recommended for training. It is retained for
provenance and to document the project's history. The only model trained on it —
SmolVLA V1 Misc —
does not work.
Part of Project-IRA — Interactive Robotic Arm.
Code:… See the full description on the dataset page: https://huggingface.co/datasets/Project-IRA/TPSoSe2026_Dataset_Collection_LeRobot_SO101.Hand_Tools
ABC-iRobotics/Hand_Tools: dataset 0000
Metric RGB-D scene dataset generated with the Label Factory workflow.
The files for this capture are stored below the 0000/ directory so
multiple numbered datasets can coexist in this repository.
Training configurations
depth_estimation: RGB input, metric depth target, intrinsics and depth units
instance_segmentation: RGB input, instance-mask target, boxes and annotations
object_pose_estimation: RGB-D input, masks, camera… See the full description on the dataset page: https://huggingface.co/datasets/ABC-iRobotics/Hand_Tools.iris-DINO-datasetIranianCarsNumberPlateIR-500KDMR-IR
Dataset Card for DMR-IR
DMR-IR is an infrared imaging dataset for Mamma Research, featuring TIFF raw temperature images and clinical metadata with generated text prompts from data acquired at Antônio Pedro University Hospital. It supports multimodal research in breast imaging and diagnostics.
Dataset Details
Dataset Description
The DMR-IR dataset is an infrared imaging database for Mamma Research, developed from clinical data acquired at the Antônio Pedro… See the full description on the dataset page: https://huggingface.co/datasets/SemilleroCV/DMR-IR.IRLBench
IRLBench: A Multi-modal, Culturally Grounded, Parallel Irish-English Benchmark for Open-Ended LLM Reasoning Evaluation
Overview
Recent advances in Large Language Models (LLMs) have demonstrated promising knowledge and reasoning abilities, yet their performance in multilingual and low-resource settings remains underexplored. Existing benchmarks often exhibit cultural bias, restrict evaluation to text-only, rely on multiple-choice formats, and, more importantly, are… See the full description on the dataset page: https://huggingface.co/datasets/ReliableAI/IRLBench.Gaze-Co-Benchmark
Gaze-Co Benchmark
Paper: Gaze Target Estimation Anywhere with Concepts
Evaluation set for Promptable Gaze Target Estimation (PGE), the task introduced in
Gaze Target Estimation Anywhere with Concepts (CVPR 2026).
Each instance pairs an image with a natural-language description of one person and the
gaze target that person is looking at. A model receives the image and the text prompt — no
head bounding box, no detector, no pose — and must localize the subject and predict where… See the full description on the dataset page: https://huggingface.co/datasets/IrohXu/Gaze-Co-Benchmark.auto-city-research
Auto-City-Research - Damage Is Not Need
Auto-City-Research
Code:
github.com/Ireliya/auto-city-research
Data:
huggingface.co/datasets/Ireliya/auto-city-research
Project:
ireliya.github.io/auto-city-research
This repository is the lightweight, privacy-safe reproducibility dataset for the Urban Cup 2026 Competition 2 project:
Damage Is Not Need: Auditing Post-Disaster Priority Disagreement with Multi-Source Urban Evidence
Research Scope… See the full description on the dataset page: https://huggingface.co/datasets/Ireliya/auto-city-research.H3-IR
H3-IR
H3-IR contains privacy-reviewed prompt/Context-IR pairs for training H3 prompt
enhancers. The public export is fail-closed: a row is included only when its
text, annotation, and every referenced media asset pass both privacy and
redistribution-rights gates.
Splits
Split
Rows
train
1110
validation
81
total
1191
Privacy Review
All source rows and unique visual assets were reviewed with gpt-5.6-sol at
reasoning_effort=xhigh… See the full description on the dataset page: https://huggingface.co/datasets/StellarVoyager/H3-IR.
