datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
exp-temporal-stability
Experiment H13: Temporal Stability Across Model Versions
Paper DOI: 10.5281/zenodo.19422427 — R15 (Zharnikov, 2026v)
Dataset DOI: 10.57967/hf/8455
Source Code: spectralbranding/sbt-papers/r15-ai-search-metamerism
Dataset Summary
450 LLM API calls testing whether successive model versions produce significantly different dimensional weight profiles for the same brands. Supplementary to the R15 study on dimensional collapse in AI-mediated brand perception (Zharnikov… See the full description on the dataset page: https://huggingface.co/datasets/spectralbranding/exp-temporal-stability.pg19-stability-bench
PG-19 Stability Evaluation Prompts
Long-context prompts in chat-message format at five bucket sizes (8K, 16K, 32K, 64K, 128K user-message tokens), designed for output-stability evaluation of LLMs under stress: as context grows (and RoPE scaling extends the effective window), do generations stay coherent — or do they degrade into mojibake, token soup, phrase loops, or script drift?
Each prompt is a 2-turn conversation (system + user) ready to send to any OpenAI-compatible… See the full description on the dataset page: https://huggingface.co/datasets/nnilayy/pg19-stability-bench.
