datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
medgemma-4b-hematologic-oncology-blind-spots
MedGemma Blind Spots: Hematologic Oncology & CAR-T Immunotherapy
A 13-probe red-team evaluation showing how Google's MedGemma-4B confidently hallucinates clinical-trial statistics, fabricates non-existent treatment regimens, and misdiagnoses lymphoma in hematologic oncology — a clinical domain absent from its documented training data.
Summary
This dataset documents failures of Google's MedGemma-4B on hematologic oncology prompts — a clinical subspecialty absent… See the full description on the dataset page: https://huggingface.co/datasets/Mateenah/medgemma-4b-hematologic-oncology-blind-spots.qwen35_4b_blindspots
Blind Spots of Qwen3.5-4B-Base
10 cases where Qwen/Qwen3.5-4B-Base gets things wrong.
Model
Name: Qwen/Qwen3.5-4B-Base
Type: pretrained base model (not instruction-tuned)
Params: ~4B
Architecture: hybrid Gated DeltaNet + Gated Attention
Released: March 2, 2026
Context: 262k tokens
Languages: 201
The model card says it's intended for "fine-tuning, in-context learning experiments, and other research or development purposes, not direct interaction."
How I… See the full description on the dataset page: https://huggingface.co/datasets/aliraza9/qwen35_4b_blindspots.
