prism-b
Datasets
All datasets matching “prism-b”PRISM_Benchmark
PRISM: PhotoRealistic Image Synthesis and Manipulation
Dataset repository for the paper "The PRISM benchmark: PhotoRealistic Image Synthesis and Manipulation to detect generated images" — Bartolucci, Salti, Lisanti.
Real images
The corresponding real images can be downloaded separately from COCO:
Split
Source
Test set
COCO 2017 Val
Training set
COCO 2017 Train
Citation
@article{BARTOLUCCI2026104826,
title = {The PRISM… See the full description on the dataset page: https://huggingface.co/datasets/oppiliF/PRISM_Benchmark.PRISM-Benchprism-boundary-bpe16k-corpusmdh-bench_multi-domain-hallucination-benchmark
MDH-Bench: Multi-Domain Hallucination Benchmark
Dataset Overview
This dataset is introduced and described in the paper:
A Unified Multi-Domain Framework for Hallucination Detection and Reliability Evaluation in Large Language Models (MDH-Bench).
The paper presents the dataset construction methodology, annotation framework, and evaluation protocols. Please refer to the paper for detailed technical descriptions.
MDH-Bench (Multi-Domain Hallucination Benchmark) is a curated… See the full description on the dataset page: https://huggingface.co/datasets/Prism-BMSCE/mdh-bench_multi-domain-hallucination-benchmark.cAI-Prism-B.5-K50
CompactAI-Prism B.5 K50
High-Density Distillation Dataset for Small Model English Language Acquisition
License: MITTop-K: 50 (Current release: K50)Source Model: Qwen3.5 2BPrimary Objective: Teach small-scale AI models to generate fluent, coherent English text through probability-aware distillation. Or at least help them sound less like they learned English from a fortune cookie.
Overview
CompactAI-Prism is a specialized training dataset designed… See the full description on the dataset page: https://huggingface.co/datasets/Glint-Research/cAI-Prism-B.5-K50.prism-bench
PRISM-Bench: Measuring Value, Evidence, and Source Hierarchies in Frontier AI Systems
Anonymous submission to NeurIPS 2026 Evaluations & Datasets Track.
License: CC BY 4.0 | Croissant: included with full RAI fields
Dataset Summary
PRISM-Bench is the first multi-model forced-choice benchmark measuring the upper three layers of the Authority Stack model:
Value (V) — which value priorities guided the decision (Schwartz 10-value framework)
Evidence (E) — which evidence type… See the full description on the dataset page: https://huggingface.co/datasets/enoch6101/prism-bench.
