datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
video-quality-scored
Image-to-Video Quality-Scored Clips
A collection of prompted image-to-video samples with quality-evaluation metadata.
Each sample pairs a first frame (the I2V conditioning image) with one or both
of:
a generated video produced by a video model from the first frame + prompt
an original clip (the reference/source video the prompt was authored around)
A subset of the samples also carry per-clip quality scores: an overall
quality_score, six per-aspect breakdowns… See the full description on the dataset page: https://huggingface.co/datasets/mohantesting/video-quality-scored.wikitext-103-quality-scored
WikiText-103 Quality-Scored Corpus
Cleaned and quality-scored subset of WikiText-103 (curated Wikipedia Good and Featured articles), prepared for character-level language model training with curriculum learning support.
Dataset Description
This dataset contains cleaned text from WikiText-103, with each batch scored on multiple quality dimensions for curriculum-based training. The text has been lowercased and filtered to an ASCII character set suitable for character-level… See the full description on the dataset page: https://huggingface.co/datasets/LisaMegaWatts/wikitext-103-quality-scored.DaTikZ-V4-quality-scored
