datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
NPM-Artifact-Explanation-Benchmark
NPM-Artifact-Explanation-Benchmark
English
NPM-Artifact-Explanation-Benchmark is a cross-category multimodal corpus and benchmark resource for Chinese cultural artifact understanding and explanation.
This release contains 28,826 cleaned artifact records derived from National Palace Museum source records' opendata (https://digitalarchive.npm.gov.tw/opendata/). Each record includes structured artifact metadata, image URLs, source record URLs, and human-written… See the full description on the dataset page: https://huggingface.co/datasets/shunanhe/NPM-Artifact-Explanation-Benchmark.eleusis-calibrated-rules
Eleusis Calibrated Rules — 100-turn reward calibration
A calibrated rule dataset for the single-player Eleusis inductive-reasoning
environment. It extends the 26-rule Hugging Face benchmark with controlled
static, transition, conditional, periodic, chunk, higher-order history, global
history, and compositional rule families.
Source benchmark: Hugging Face Eleusis.
Dataset version: v2.1-frontier-calibrated-100turn-20260812Protocol: eleusis-100-v11
The structural, GPT Sol… See the full description on the dataset page: https://huggingface.co/datasets/nph4rd/eleusis-calibrated-rules.MS-GPT-NPLIB1
MS-GPT NPLIB1 benchmark data
This repository contains the NPLIB1 benchmark files released with
MS-GPT: Rethinking MS/MS De Novo Structure Elucidation as Spectrum-Induced
Posterior Querying of a Molecule-Language Model.
Provenance and attribution
These files are mirrored from the authors' released
MS-GPT asset bundle
with their explicit authorization.
Original repository: https://github.com/VIKI623/MS-GPT
Author's Hugging Face profile:… See the full description on the dataset page: https://huggingface.co/datasets/nielsr/MS-GPT-NPLIB1.
