datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SteamScreenshots-Bugs
Samples
arxiv_dbarxiv_qa2CHMCorrUserResponseimagenet_hard_review_data_r2GBPro-RAWVideoGameQA-Bench
VideoGameQA-Bench: Evaluating Vision-Language Models for Video Game Quality Assurance
by Mohammad Reza Taesiri, Abhijay Ghildyal, Saman Zadtootaghaj, Nabajeet Barman, Cor-Paul Bezemer
Abstract:
With video games now generating the highest revenues in the entertainment industry, optimizing game development workflows has become essential for the sector's sustained growth. Recent advancements in Vision-Language Models (VLMs) offer considerable potential to automate and… See the full description on the dataset page: https://huggingface.co/datasets/taesiri/VideoGameQA-Bench.steam_screenshots_samples_2arxiv_audioArXivSignals-DeepSummaries
ArXivSignals DeepSummaries — Agent-Built Visual Paper Explainers
A continuously-updated, day-partitioned dataset of deep, visual summaries of
arXiv papers, each built by a coding agent working inside the paper's own
LaTeX source: the agent reads the full text, authors an editorial narrative as
a structured content spec, and the paper's real figures and tables
(extracted and rendered from the LaTeX, web-optimized) ride along as an
embedded, variable-length image array. The… See the full description on the dataset page: https://huggingface.co/datasets/taesiri/ArXivSignals-DeepSummaries.imagenet-hard-4K
Dataset Card for "Imagenet-Hard-4K"
Project Page - Paper - Github
ImageNet-Hard-4K is 4K version of the original ImageNet-Hard dataset, which is a new benchmark that comprises 10,980 images collected from various existing ImageNet-scale benchmarks (ImageNet, ImageNet-V2, ImageNet-Sketch, ImageNet-C, ImageNet-R, ImageNet-ReaL, ImageNet-A, and ObjectNet). This dataset poses a significant challenge to state-of-the-art vision models as merely zooming in often fails to improve their… See the full description on the dataset page: https://huggingface.co/datasets/taesiri/imagenet-hard-4K.FragileXSteamScreenshotsArXivSignals
ArXivSignals — Daily arXiv Papers with LLM Signal & Summaries
A continuously-updated, day-partitioned dataset of arXiv papers (AI/ML and
adjacent categories) enriched with LLM-derived signal: a 0–100 importance
score, topical/lab tags, a one-line takeaway, and — for a selected subset —
dense full-page summaries. It powers arxivsignals.io
and is published here as an open research resource.
The dataset has two configs:
papers (default) — one row per paper: bibliography +… See the full description on the dataset page: https://huggingface.co/datasets/taesiri/ArXivSignals.imagenet_hard_review_dataKo-StrategyQA
Ko-StrategyQA
This dataset represents a conversion of the Ko-StrategyQA dataset into the BeIR format, making it compatible for use with mteb.
The original dataset was designed for multi-hop QA, so we processed the data accordingly. First, we grouped the evidence documents tagged by annotators into sets, and excluded unit questions containing 'no_evidence' or 'operation'.
egoxtreme
EgoXtreme: A Dataset for Robust Object Pose Estimation in Egocentric Views under Extreme Conditions
📖 Dataset Information
EgoXtreme is a novel large-scale dataset designed for robust egocentric 6D object pose estimation under extreme environmental conditions. The dataset comprises approximately 1.3 million frames with a total duration of 775.5 minutes (~12.9 hours). It was captured at 30 fps using Aria glasses, providing high-resolution 1408 x 1408 raw fisheye RGB… See the full description on the dataset page: https://huggingface.co/datasets/taegyoun88/egoxtreme.arxiv_summaryacoustic_context_switchingmsr-acc-tae25
Microsoft Research - Accurate Chemistry Collection: Total Atomization Energies
Description
The Microsoft Research Accurate Chemistry Collection (MSR-ACC) provides a collection of accurate coupled cluster labels for training machine learning functionals.
MSR-ACC/TAE25 comprising 73,040 total atomization energies at the CCSD(T)/CBS level obtained with the W1-F12 thermochemical protocol.
The dataset is constructed to exhaustively cover the chemical space of closed-shell… See the full description on the dataset page: https://huggingface.co/datasets/microsoft/msr-acc-tae25.tae-data-split-paragraphs
Split Paragraphs Dataset
Split paragraphs data with configs 000-099.
GameplayCaptions-GPT-4VGameplayCaptions-Gemini-pro-visionSteamGlitches-Gemini-Labelstae-data-embeddingsarxiv_qa
ArXiv QA
(TBD) Automated ArXiv question answering via large language models
Github | Homepage | Simple QA - Hugging Face Space
Automated Question Answering with ArXiv Papers
Latest 25 Papers
LIME: Localized Image Editing via Attention Regularization in Diffusion
Models - [Arxiv] [QA]
Revisiting Depth Completion from a Stereo Matching Perspective for
Cross-domain Generalization - [Arxiv] [QA]
VL-GPT: A Generative Pre-trained Transformer for Vision and… See the full description on the dataset page: https://huggingface.co/datasets/taesiri/arxiv_qa.kala-uap-archiveimagenet-hard
Dataset Card for "ImageNet-Hard"
Project Page - ArXiv - Paper - Github - Image Browser
Dataset Summary
ImageNet-Hard is a new benchmark that comprises 10,980 images collected from various existing ImageNet-scale benchmarks (ImageNet, ImageNet-V2, ImageNet-Sketch, ImageNet-C, ImageNet-R, ImageNet-ReaL, ImageNet-A, and ObjectNet). This dataset poses a significant challenge to state-of-the-art vision models as merely zooming in often fails to improve their ability to… See the full description on the dataset page: https://huggingface.co/datasets/taesiri/imagenet-hard.GameplayCaptions-GPT-4V-V2Gameplay-Walkthrough-QA
