CoolFace
12 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sjtu-deepvision /DRR_data0 likes5k downloads10mo agoHugging Face02skylenage-ai /DeepVision-103K 🔭 DeepVision-103K A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning Training on DeepVision-103K yields top performance on both multimodal mathematical reasoning and general multimodal benchmarks: Average Performance on multimodal math and general multimodal benchmarks. Training on DeepVision-103K elicits more efficient reasoning. Benchmark Qwen3-VL-8B-Instruct (Acc / Tokens) Qwen3-VL-8B-DeepVision (Acc /… See the full description on the dataset page: https://huggingface.co/datasets/skylenage-ai/DeepVision-103K.imageimage-text-to-text100K<n<1M36 likes972 downloads7mo agoHugging Face03Devilishcode /DeepVision-103K 🔭 DeepVision-103K A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning Training on DeepVision-103K yields top performance on both multimodal mathematical reasoning and general multimodal benchmarks: Average Performance on multimodal math and general multimodal benchmarks. Training on DeepVision-103K elicits more efficient reasoning. Benchmark Qwen3-VL-8B-Instruct (Acc / Tokens) Qwen3-VL-8B-DeepVision (Acc /… See the full description on the dataset page: https://huggingface.co/datasets/Devilishcode/DeepVision-103K.imageimage-text-to-text100K<n<1M0 likes79 downloads7mo agoHugging Face04blsmash044 /DeepVision-103K 🔭 DeepVision-103K A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning Training on DeepVision-103K yields top performance on both multimodal mathematical reasoning and general multimodal benchmarks: Average Performance on multimodal math and general multimodal benchmarks. Training on DeepVision-103K elicits more efficient reasoning. Benchmark Qwen3-VL-8B-Instruct (Acc / Tokens) Qwen3-VL-8B-DeepVision (Acc /… See the full description on the dataset page: https://huggingface.co/datasets/blsmash044/DeepVision-103K.imageimage-text-to-text100K<n<1M0 likes58 downloads7mo agoHugging Face05JamesGoGo /DeepVision-103K 🔭 DeepVision-103K A Visually Diverse, Broad-Coverage, and Verifiable Mathematical Dataset for Multimodal Reasoning Training on DeepVision-103K yields top performance on both multimodal mathematical reasoning and general multimodal benchmarks: Average Performance on multimodal math and general multimodal benchmarks. Training on DeepVision-103K elicits more efficient reasoning. Benchmark Qwen3-VL-8B-Instruct (Acc / Tokens) Qwen3-VL-8B-DeepVision (Acc /… See the full description on the dataset page: https://huggingface.co/datasets/JamesGoGo/DeepVision-103K.imageimage-text-to-text100K<n<1M0 likes56 downloads7mo agoHugging Face06williamhkml /deepvision_datasetsvideon<1K0 likes52 downloads28d agoHugging Face07SanjanaMaria /deepvision-vlm-predictions VLM Evaluation Predictions — Zeroshot-DeepVision-24k Model evaluation predictions and metrics from the Bengali Math VQA pipeline. Folder structure Baseline evaluation protocol Base model output is constrained via a system-level instruction to output only the Bengali MCQ option letter (ক/খ/গ/ঘ), ensuring format parity with the fine-tuned model. visual-question-answering0 likes16 downloads5mo agoHugging Face08RaiyanKhaan /deepvision-zero-shot-20kimage10K<n<100K0 likes13 downloads6mo agoHugging Face09UE-APC-INRAE /deepvisiontools-demo-datasets Description Those are 3 datasets for deepvisiontools library demo. deepvisiontools homepage : https://forge.inrae.fr/ue-apc/librairies/python/deepvisiontools Datasets The dice dataset was downloaded from Kaggle : https://www.kaggle.com/datasets/nellbyler/d6-dice The VegannSubDataset is a small portion from : https://zenodo.org/records/7636408 The coco_6cls_subset was obtained from : https://universe.roboflow.com/nan-ixwz3/coco-y1tdb image1K<n<10K0 likes5 downloads5mo agoHugging Face10UE-APC-INRAE /deepvisiontools_tutorials0 likes4 downloads4mo agoHugging Face11SanjanaMaria /DeepVision_ZS_Predictions VLM Evaluation Predictions Prediction outputs from the Zeroshot-DeepVision-24k Bengali Math VQA evaluation pipeline. Structure {model_short}/ baseline-testing/ ← predictions BEFORE fine-tuning post-finetune-testing/ ← predictions AFTER fine-tuning Generated by the reusable Kaggle evaluation notebook. tabularn<1K0 likes2 downloads5mo agoHugging Face12davanstrien /deepvision-atlas-data0 likes1 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.