datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
smollm3-3b-base-blindspots
SmolLM3-3B-Base — Blind Spots Dataset
This dataset contains 10 diverse input-output pairs where the base language model
HuggingFaceTB/SmolLM3-3B-Base
produces incorrect predictions under greedy decoding. Each row records the exact prompt fed
to the model, the correct expected answer, and what the model actually generated — along with
a description of the error type.
Model Tested
Field
Value
Model
HuggingFaceTB/SmolLM3-3B-Base
Parameters
3 billion… See the full description on the dataset page: https://huggingface.co/datasets/Dhruba461/smollm3-3b-base-blindspots.smollm3-3b-base-blind-spots
Blind Spots of SmolLM3-3B-Base
This dataset contains 15 diverse examples where the base language model HuggingFaceTB/SmolLM3-3B-Base produces incorrect, repetitive, or off‑task outputs. It was created as part of the Fatima Fellowship technical challenge.
Model
Name: HuggingFaceTB/SmolLM3-3B-Base
Type: Decoder‑only transformer (base model, not instruction‑tuned)
Parameters: 3B
Link: https://huggingface.co/HuggingFaceTB/SmolLM3-3B-Base
Methodology
I loaded the… See the full description on the dataset page: https://huggingface.co/datasets/AyshSaleem/smollm3-3b-base-blind-spots.
