CoolFace
Datasetpublic

ALb78/qwen2_5_reasoning_failures

Reasoning and Logic Failure Cases in Qwen2.5-1.5B Diagnostic dataset of reasoning errors in a small base language model Technical challenge: Blind Spots of Frontier Models by Fatima Institute for Global AI Research Overview This dataset documents systematic reasoning failures observed while evaluating the base language model Qwen/Qwen2.5-1.5B. The dataset records cases where the model produces confident but incorrect answers to questions requiring:… See the full description on the dataset page: https://huggingface.co/datasets/ALb78/qwen2_5_reasoning_failures.

sourceHugging Faceupdated 7mo agoView on Hugging Face
0likes2downloads
settings

This repository belongs to ALb78 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameqwen2_5_reasoning_failures
visibilitypublic
licencenot set
gatedno
ownerALb78
Account settings
ALb78/qwen2_5_reasoning_failures · CoolFace