literary
Datasets
All datasets matching “literary”literary-analogy
NuBerea/literary-analogy
An access-controlled registry of attested literary analogies and narrative
parallels in the Hebrew Bible. It represents parallel loci, their passage
members and spans, structured scholarly attestations, and an explicitly
provisional verse-alignment surface.
The registered NuBerea tools expose four governed query views: passage profiles,
pair alignments, parallel-family browsing, and attestation provenance. Internal
bookkeeping and annotation configs are… See the full description on the dataset page: https://huggingface.co/datasets/NuBerea/literary-analogy.LiteraryQALiteraryQA is a dataset for question answering over narrative text, specifically books. It is a cleaned subset of the NarrativeQA dataset, focusing on books from Project Gutenberg with improved text quality and formatting and better question-answer pairs.gallica_literary_fictions
Dataset Card for Literary fictions of Gallica
Dataset Summary
The collection "Fiction littéraire de Gallica" includes 19,240 public domain documents from the digital platform of the French National Library that were originally classified as novels or, more broadly, as literary fiction in prose. It consists of 372 tables of data in tsv format for each year of publication from 1600 to 1996 (all the missing years are in the 17th and 20th centuries). Each table is… See the full description on the dataset page: https://huggingface.co/datasets/biglam/gallica_literary_fictions.JP-TH_Literary_Translation_URL_Alignment_Index
JP–TH Literary Translation URL Alignment Index
This release provides a copyright-conscious metadata index and reproducibility package for a Japanese–Thai literary translation dataset associated with the study Context-Aware Prompting for Japanese–Thai Literary Translation in a Low-Resource Setting.
Overview
The release is designed to support reproducible academic research on Japanese–Thai literary machine translation, context-aware prompting, prompt engineering… See the full description on the dataset page: https://huggingface.co/datasets/Gsk068/JP-TH_Literary_Translation_URL_Alignment_Index.literary-roleplay
Dataset Card for Literary Roleplay SFT
Dataset Summary
An instruction-tuning dataset for training models to roleplay properly, derived from literary sources across five languages and three roleplay-engine logic frameworks. The dataset contains 346 rows spanning English (164), Russian (68), Hindi (38), Sanskrit (38), and Japanese (38), drawn from the works of Gogol, Bulgakov, Perumov, Golovachev, Vedic canon (Upanishads, Mahabharata, Ramayana), classic sci-fi… See the full description on the dataset page: https://huggingface.co/datasets/Exxe/literary-roleplay.fr_literary_dataset_base
