fiction
Datasets
All datasets matching “fiction”fictionrag-datasetfictional-characters-image-datasetHow to use
Here is how to use this dataset:
from datasets import load_dataset
dataset = load_dataset("gryffindor-ISWS/fictional-characters-image-dataset")
This repository contains fictional characters dataset constructed from Wikidata for the research project "Draw Me Like Your Triples: Leveraging Generative AI for the Completion of Wikidata". The project was conducted by Raia Abu Ahmad, Martin Critelli, Şefika Efeoğlu, Eleonora Mancini, Célian Ringwald and Xinyue Zhang under the… See the full description on the dataset page: https://huggingface.co/datasets/gryffindor-ISWS/fictional-characters-image-dataset.VellumK2T-Fiction-SFT-01
Dataset Card for VellumK2T-Fiction-SFT-01
A long-form synthetic creative fiction dataset with 8,042 instruction–output pairs for supervised fine-tuning (SFT), generated using the VellumForge2 pipeline and published as part of the VellumForge2 fantasy collection on Hugging Face.
Dataset Details
Dataset Description
VellumK2T-Fiction-SFT-01 is a synthetically generated dataset of various fiction writing samples. Each row contains:
An instruction: a rich… See the full description on the dataset page: https://huggingface.co/datasets/lemon07r/VellumK2T-Fiction-SFT-01.2026-08-27-good-ai-fiction-sf-860
synth good_ai_fiction run — per-stage snapshots (resumable generation cache)
field
value
experiment
synth good_ai_fiction run — per-stage snapshots (resumable generation cache)
date_generated
20260828_020624
constitution
constitutions/claude_distilled_12_principles_mid/constitution.md
source_repo
https://github.com/Matthew-Bozoukov/teaching_claude_why_replication.git @ ae0725130a2fccd74fe7bdef5c570ec71420cd7b
models
per-stage models — see manifest.json… See the full description on the dataset page: https://huggingface.co/datasets/dougalldeepmind/2026-08-27-good-ai-fiction-sf-860.fictionalqa_training_splits
Training splits view of the FictionalQA dataset
The FictionalQA dataset
Repository: https://github.com/jwkirchenbauer/fictionalqa
Paper: https://arxiv.org/abs/2506.05639
Dataset Description
This dataset is a derivative of the main dataset hf.co/datasets/jwkirchenbauer/fictionalqa. Please see that dataset's README for a detailed description of the assets.
The dataset splits (configs) provided here are the exact ones materialized and used in the experiments for… See the full description on the dataset page: https://huggingface.co/datasets/jwkirchenbauer/fictionalqa_training_splits.gallica_literary_fictions
Dataset Card for Literary fictions of Gallica
Dataset Summary
The collection "Fiction littéraire de Gallica" includes 19,240 public domain documents from the digital platform of the French National Library that were originally classified as novels or, more broadly, as literary fiction in prose. It consists of 372 tables of data in tsv format for each year of publication from 1600 to 1996 (all the missing years are in the 17th and 20th centuries). Each table is… See the full description on the dataset page: https://huggingface.co/datasets/biglam/gallica_literary_fictions.
