datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dactylic-hexameter-latin-poetry-corpus
Dactylic Hexameter Latin Poetry Corpus
This repository contains a curated and processed corpus of Classical Latin poetry written in dactylic hexameter. It serves as the raw training data ("Dataset V3") for the Master's Thesis titled "A Hybrid Post Hoc Feedback Framework for Latin Dactylic Hexameter" submitted to KU Leuven (2025).
Dataset Description
This corpus was constructed to fine-tune Large Language Models (LLMs) for the generation of metrically valid Latin poetry.… See the full description on the dataset page: https://huggingface.co/datasets/KaanGoker/dactylic-hexameter-latin-poetry-corpus.SupraReviewBench
SupraReviewBench
Dataset summary
SupraReviewBench is a peer-review benchmark built from OpenReview discussion threads.
Each record represents one paper and its full review discussion. Reviewer opinions are
split into atomic blocks, labeled with a taxonomy, grouped by discussion point, and
validated for correctness via conflict adjudication and author-refutation analysis.
The dataset is intended for opinion-level evaluation and training, with explicit
labels that mark… See the full description on the dataset page: https://huggingface.co/datasets/hexuandeng/SupraReviewBench.
