datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
catholic-resources
Vietnamese Catholic resources by v-bible
Data Structure
calendar: Generated Liturgical calendars using
v-bible/js-sdk.
misc/proper-names.json: Name translation from
ktcgkpv.org, generated by
v-bible/bible-scraper.
liturgical: Liturgical data from
The Lectionary for Mass (1998/2002 USA Edition),
compiled by Felix Just, S.J., Ph.D., and generated by
v-bible/bible-scraper.
books/bible: Generated Bible markdown data.
books/catechism-books: Official catechism… See the full description on the dataset page: https://huggingface.co/datasets/v-bible/catholic-resources.OmniRooms
UniSHARP:
Universal Sharp Monocular View Synthesis
Meixi Song1 ·
Dizhe Zhang1,* ·
Hao Ren1 ·
Ruiyang Zhang1 ·
Bo Du2 ·
Ming-Hsuan Yang3 ·
Lu Qi1,2,*
1Insta360 Research · 2Wuhan University · 3University of California, Merced
UniSHARP extends SHARP-style photorealistic monocular view synthesis to universal camera systems. Given a single image from a perspective, wide-FoV, fisheye, or panoramic camera, UniSHARP predicts a 3D Gaussian representation and… See the full description on the dataset page: https://huggingface.co/datasets/Insta360-Research/OmniRooms.gaia2_filesystem
GAIA2 Filesystem
This is a dataset containing files for the GAIA2 benchmark. You should not use this dataset on its own, but instead use the Meta Agents Research Environments framework to execute scenarios from that GAIA2 dataset.
Dataset Link
https://huggingface.co/datasets/meta-agents-research-environments/gaia2
Contact Details
Publishing POC: Meta AI Research Team
Affiliation: Meta Platforms, Inc.
Website:… See the full description on the dataset page: https://huggingface.co/datasets/meta-agents-research-environments/gaia2_filesystem.imagenet_1k_resized_256
Dataset Card for "imagenet_1k_resized_256"
Dataset summary
The same ImageNet dataset but all the smaller side resized to 256.
A lot of pretraining workflows contain resizing images to 256 and random cropping to 224x224, this is why 256 is chosen.
The resized dataset can also be downloaded much faster and consume less space than the original one.
See here for detailed readme.
Dataset Structure
Below is the example of one row of data. Note that the labels in… See the full description on the dataset page: https://huggingface.co/datasets/evanarlian/imagenet_1k_resized_256.kaz-vision-50kD3HRconceptual_captions
Dataset Card for Conceptual Captions
Dataset Summary
Conceptual Captions is a dataset consisting of ~3.3M images annotated with captions. In contrast with the curated style of other image caption annotations, Conceptual Caption images and their raw descriptions are harvested from the web, and therefore represent a wider variety of styles. More precisely, the raw descriptions are harvested from the Alt-text HTML attribute associated with web images. To arrive at the… See the full description on the dataset page: https://huggingface.co/datasets/google-research-datasets/conceptual_captions.phaseShift_shell_result_pdf
Phase Resonance / IRS-DCE
Topological Dynamics & Artificial Cognitive Physics
Open Structural Record of Basis-Relative Reorganization in Transformer Representation Space
All pdf Creative Creative Commons Attribution No Derivatives 4.0 International
This repository provides comprehensive PDF research materials and Python scripts and something for mathematical proofs in the field of AI.
📢 Achievement: total 51k+ Downloads!… See the full description on the dataset page: https://huggingface.co/datasets/meta13sphere/phaseShift_shell_result_pdf.resisc45
RESISC45
Overview
Usage
from datasets import load_dataset
# Load the dataset
dataset = load_dataset('tanganke/resisc45')
Dataset Information
The dataset is divided into the following splits:
Training set: Contains 18,900 examples, used for model training.
Test set: Contains 6,300 examples, used for model evaluation and benchmarking.
The dataset also includes the following augmented sets, which can be used for testing the model's robustness to… See the full description on the dataset page: https://huggingface.co/datasets/tanganke/resisc45.showdown-shower-resourcesResources for Showdown Shower
OmegaUse-OfficeVal
OmegaUse-OfficeVal
Benchmarking LLM Agents on Long-Horizon Office-Suite Tasks with Economic Grounding
OmegaUse-OfficeVal is a benchmark for evaluating LLM agents on long-horizon,
real-world office-suite tasks that span word-processing documents, spreadsheets,
presentations, and cross-file productivity workflows. Tasks are derived from
authentic office requests proposed by practitioners and drawn from freelance
platforms, grounding the benchmark in real economic demand. Each task… See the full description on the dataset page: https://huggingface.co/datasets/baidu-frontier-research/OmegaUse-OfficeVal.resisc45
Description
RESISC45 dataset is a publicly available benchmark for Remote Sensing Image Scene Classification (RESISC), created by Northwestern Polytechnical University (NWPU). This dataset contains 31,500 images, covering 45 scene classes with 700 images in each class.
The dataset does not have any default splits. Train, validation, and test splits were based on these definitions here… See the full description on the dataset page: https://huggingface.co/datasets/timm/resisc45.sit-experiment-results-2026-04-16
SiT Experiment Results
This dataset repo contains experiment outputs generated from the SiT-XL/2 model under the locally adjusted protocol in gpt_experiment_instructions.md.
Protocol Summary
Main dataset size: 20 images
Control dataset size: 10 images
Main/control images were exported from timm/mini-imagenet
Model: SiT-XL/2
VAE: stabilityai/sd-vae-ft-ema
Included Results
task0_outputs/: Prompt 0 / Task 0 artifacts
task1_outputs_20/: basic statistics… See the full description on the dataset page: https://huggingface.co/datasets/LamTNguyen/sit-experiment-results-2026-04-16.relaion2B-en-researchKohaku-Delta-Alpha-Characters-Result
Test Index for Kohaku-Delta-Alpha Model
ID
Tag
Copyright
Gender
Posts
CCIP
AIC
BP
Core Tags
2075918
inkling_player_character
splatoon_(series)
female
9646
0.383333
0.998157
0.645165
inkling girl, long hair, tentacle hair, pointy ears, bangs, blunt bangs, red eyes
2040387
doodle_sensei_(blue_archive)
blue_archive
female
6632
0.721212
0.992072
0.625653
halo, bangs, blue eyes, breasts, long hair, black hair, blue hair, hair ornament
1978860
2b_(nier:automata)… See the full description on the dataset page: https://huggingface.co/datasets/AngelBottomless/Kohaku-Delta-Alpha-Characters-Result.nine-source-hand-data-review-results
九源手部数据:修正版文件夹交付
新版共验收通过 457 个完整彩色 MANO 双栏视频,公开 138 个;其余明确列为未完成或诊断。旧版骨架视频已从最终 demo 展示撤下。
新版视频文件夹 · 逐源验收及未完成项 · 旧版诊断区 · 分布、benchmark、结论 · 筛选清单 · 完整文件表
来源
上下文
已渲染
彩色完整 demo 通过
公开新版
未通过/未完成视频
H2O
50
50
50
0
0
HOT3D
100
200
58
58
142
HOI4D
50
50
50
0
0
ARCTIC
100
200
173
0
27
Ego4D
0
0
0
0
0
EgoDex
50
50
46
0
4
EPIC-KITCHENS
50
49
30
30
20
HO3D
0
0
0
0
0
DexYCB
50
50
50
50
0
Ego4D/HO3D 本批各 50 个目标片段未完成。HO3D 下载已复查可用,详见 访问复查。
hf download… See the full description on the dataset page: https://huggingface.co/datasets/yangzijing/nine-source-hand-data-review-results.GenIRspatial-moe-resultsbankertoolbench
BankerToolBench
BankerToolBench is a benchmark of 100 end-to-end investment banking tasks for
evaluating AI agents. Each task mirrors real junior-banker work — building
financial models, preparing pitch decks, writing memos — and produces multi-file
deliverables (Excel, PowerPoint, Word) that are scored against expert-authored
rubrics.
The benchmark was developed with 502 investment bankers from firms including
Goldman Sachs, JPMorgan, Evercore, and others. Human completion time… See the full description on the dataset page: https://huggingface.co/datasets/handshake-ai-research/bankertoolbench.lgg-mri-segmentation-research
LGG Brain MRI Segmentation with Genomic Clusters
This repository provides a Patient-Centric version of the Lower-Grade Glioma (LGG) Segmentation dataset. While other versions of this data exist, they often treat slices as independent images. This version preserves the 3D patient volume and integrates all genomic/clinical labels directly into a multimodal-ready format.
🌟 Why This Version?
Developed for Multimodal AI Research, this dataset addresses several limitations… See the full description on the dataset page: https://huggingface.co/datasets/Ehsan-rmz/lgg-mri-segmentation-research.AuraFusion360_ResultsDefactify_Image_Dataset
Defactify_Image_Dataset
This dataset is associated with the paper A Comprehensive Dataset for Human vs. AI Generated Image Detection.
📝 Dataset Description
Dataset Summary
The Defactify_Image_Dataset (A Comprehensive Dataset for Human vs. AI Generated Image Detection) is a high-quality collection of 96,000 images and associated metadata designed to benchmark models for detecting and identifying the source of artificially generated content. Built using the MS… See the full description on the dataset page: https://huggingface.co/datasets/Rajarshi-Roy-research/Defactify_Image_Dataset.relaion2B-en-research-safeflickr_faces_res512_50k
Dataset Card for flicker-faces
This is a FiftyOne dataset with 52001 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
from fiftyone.utils.huggingface import load_from_hub
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset = load_from_hub("harpreetsahota/flickr_faces_res512_50k")
# Launch the App
session = fo.launch_app(dataset)… See the full description on the dataset page: https://huggingface.co/datasets/harpreetsahota/flickr_faces_res512_50k.VAREX
VAREX: A Benchmark for Multi-Modal Structured Extraction from Documents
VAREX (VARied-schema EXtraction) is a benchmark for evaluating multimodal foundation models on structured data extraction from government forms. It comprises 1,777 documents with 1,771 unique schemas across three structural categories, each provided in four input modalities. Ground truth is deterministic — generated via a Reverse Annotation pipeline that programmatically fills PDF templates with synthetic values… See the full description on the dataset page: https://huggingface.co/datasets/ibm-research/VAREX.resfit_robocasa_publicvisual-reasoning-benchmark-results
Visual Reasoning Benchmark Suite v3.3 · 2005 Tasks · 12 Tracks Equal Weight
本版本以用户最新上传的 visual_reasoning_benchmark_suite_v3_修改 为唯一基础版本,不回退、不覆盖用户已经重绘或修改过的既有数据。完整性比对结果:原基础包中 3283 个既有数据文件全部保持字节级不变。
在此基础上新增并整合:
Nonogram(数织)150 题:45 Easy / 60 Medium / 45 Hard;
Tangram(七巧板)150 题:45 Easy / 60 Medium / 45 Hard;
两个任务的一键生成器、统一生成入口、统一评估入口、雷达图和排行榜支持。
最终总规模:2005 题,12 个 Track。
任务与数量
Task
Count
figure_completion
394
spatial_generation
56
maze_beginner
64… See the full description on the dataset page: https://huggingface.co/datasets/songyiren/visual-reasoning-benchmark-results.CAD_tmp_resultsSiT-PCA-FID4K-results
SiT-B/2 single-block PCA experiments
This repository contains quantitative and qualitative results for inserting a PCA projection and inverse reconstruction after each individual transformer block of SiT-B/2. It covers all 48 combinations of:
block: 1 through 12;
rank: 64 and 128;
PCA axis: hidden and token.
Each experiment modifies exactly one block. There are no joint/multi-block interventions in this sweep.
Method
The activation entering the PCA hook has… See the full description on the dataset page: https://huggingface.co/datasets/LamTNguyen/SiT-PCA-FID4K-results.video-game-super-resolution
