datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pubmedqa-recursive-llm-degradation-qwen2.5-3b
PubMedQA Recursive LLM Degradation — Qwen2.5-3B
This repository contains synthetic biomedical question-answering data
and model predictions generated as part of a study of recursive
fine-tuning and model degradation.
Base Model
Qwen/Qwen2.5-3B
Source Dataset
The experiments use the PubMedQA dataset:
qiaoxin/PubMedQA
This repository contains generated/derived research artifacts and does
not redistribute the original PubMedQA dataset in its entirety.… See the full description on the dataset page: https://huggingface.co/datasets/chrislimbe/pubmedqa-recursive-llm-degradation-qwen2.5-3b.smollm3-3b-base-blindspots
SmolLM3-3B-Base — Blind Spots Dataset
This dataset contains 10 diverse input-output pairs where the base language model
HuggingFaceTB/SmolLM3-3B-Base
produces incorrect predictions under greedy decoding. Each row records the exact prompt fed
to the model, the correct expected answer, and what the model actually generated — along with
a description of the error type.
Model Tested
Field
Value
Model
HuggingFaceTB/SmolLM3-3B-Base
Parameters
3 billion… See the full description on the dataset page: https://huggingface.co/datasets/Dhruba461/smollm3-3b-base-blindspots.sample_form_data_for_llama3.2_3b
