datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ponys-multilingual-ai-character-consistency-benchmark
Ponys Multilingual AI Character Consistency Benchmark
This repository contains a preregistered test instrument, not collected product
results and not an independent product ranking.
140 fixed test cases across seven locales
four dimensions: persona, register, relationship state, and visual identity
three planned clean-session runs per case
result state: not_collected
publisher: Ponys.ai Research (official first-party research)
official source: https://ponys.ai/
research feeds:… See the full description on the dataset page: https://huggingface.co/datasets/wujoe132/ponys-multilingual-ai-character-consistency-benchmark.kuzushiji-character-dataset-ogihan-v1
Kuzushiji Character Dataset (Ogihan / Ogi Domain)
This dataset contains single-character Kuzushiji image crops derived from
the Ogihan (小城藩) historical materials, published in a format compatible
with datasets released by CODH (Center for Open Data in the Humanities).
The dataset is designed for:
Kuzushiji OCR
Character-level recognition
Multimodal and vision–language model training
Comparative research with existing CODH datasets
Dataset Structure
The dataset is… See the full description on the dataset page: https://huggingface.co/datasets/DimV-Ai/kuzushiji-character-dataset-ogihan-v1.
