akotet08/tiny-aya_failure
Dataset Summary A compact diagnostic benchmark for evaluating failure modes in language models. It tests a model's ability to resist false premises, avoid fabricating entities, detect contradictions, follow strict constraints, and recognize category mismatches. Each record includes: id, category, prompt, model_output, and expected_correct_output. Model Outputs were generated using CohereLabs/tiny-aya-global via the Hugging Face transformers library. Generation… See the full description on the dataset page: https://huggingface.co/datasets/akotet08/tiny-aya_failure.
This repository belongs to akotet08 on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
