datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qwen3.5-2b-base-blind-spots
Qwen3.5-2B-Base — Blind Spot Analysis (Text + Vision)
Model Tested
Field
Value
Model
Qwen/Qwen3.5-2B-Base
Parameters
2.27 B (2,274 M per HF metadata)
Architecture
Hybrid Gated-DeltaNet (dense FFN) — 24 LM layers (18 DeltaNet + 6 full-attention), ViT vision encoder
Type
Pre-trained base model (not instruction-tuned)
Context
262 144 tokens
Modalities
Text + Vision (early-fusion multimodal)
Key Contributions
Only multimodal… See the full description on the dataset page: https://huggingface.co/datasets/F555/qwen3.5-2b-base-blind-spots.SpotAgenticCoT
SpotAgenticCoT: Agentic Trajectories for Visual Geo-localization
Project Page
Dataset Description
SpotAgenticCoT (specifically referring to the SpotAgenticCoT-6k subset described in the paper) is a high-quality dataset of ~6,000 agentic reasoning trajectories designed for visual geo-localization tasks.
Unlike traditional geo-localization datasets that only provide image-coordinate pairs, SpotAgenticCoT contains full ReAct (Reasoning + Acting) traces. Each sample… See the full description on the dataset page: https://huggingface.co/datasets/jiafr1802/SpotAgenticCoT.
