datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Total-Text-DatasetTotal Text Dataset.
It consists of 1555 images with more than 3 different text orientations: Horizontal, Multi-Oriented, and Curved, one of a kind.
Original github repo; https://github.com/cs-chan/Total-Text-Dataset
Forked repo; https://github.com/yunusserhat/Total-Text-Dataset
Total-Text-Dataset
Dataset Card for Total-Text-Dataset
The Total-Text consists of 1555 images with more than 3 different text orientations: Horizontal, Multi-Oriented, and Curved
This is a FiftyOne dataset with 1555 samples.
Installation
If you haven't already, install FiftyOne:
pip install -U fiftyone
Usage
import fiftyone as fo
import fiftyone.utils.huggingface as fouh
# Load the dataset
# Note: other available arguments include 'max_samples', etc
dataset =… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/Total-Text-Dataset.cvsearch_hr8kVisualizations of CVSearch
Citation
@misc{li2026cvsearchempoweringmultimodalllms,
title={CVSearch: Empowering Multimodal LLMs with Cognitive Visual Search for High-Resolution Image Perception},
author={Liupeng Li and Haoqian Kang and Zhenyu Lu and Jinpeng Wang and Bin Chen and Ke Chen and Yaowei Wang},
year={2026},
eprint={2605.23655},
archivePrefix={arXiv},
primaryClass={cs.CV},
url={https://arxiv.org/abs/2605.23655},
}… See the full description on the dataset page: https://huggingface.co/datasets/tothanhdat/cvsearch_hr8k.TotalSegmentator-CT-Lite
About
This is a derivative of the TotalSegmentator dataset.
1228 CT images and corresponding segmentation mask of 117 structures
We combined multiple segmentation masks into a single nii.gz file under the folder Masks,
and moved all CT images to the folder Images.
All images and masks are renamed according to case IDs.
This dataset is released under the CC-BY-4.0 license.
News 🔥
[10 Oct, 2025] This dataset is integrated into 🔥MedVision🔥
Official… See the full description on the dataset page: https://huggingface.co/datasets/YongchengYAO/TotalSegmentator-CT-Lite.Totalsegmentor_Pelvis_Bone_Recon_DatasetTotalSegmentatorMR
TotalSegmentator MRI (v2.0.0)
Public release of the TotalSegmentator MRI dataset: 616 heterogeneous MRI scans (T1, T2, PD, DIXON-style and other sequences, multiple field strengths, scanners, slice thicknesses, contrast agents) with voxel-wise annotations of 50 anatomical structures spanning the whole body.
Dataset Summary
Field
Details
Modality
MRI (sequence-independent: T1w / T2w / PD / DIXON / mixed)
Body Part
Whole-body / multi-organ
Structures… See the full description on the dataset page: https://huggingface.co/datasets/MedOtter/TotalSegmentatorMR.armbench-segmentation-mix-object-toteThis is data from the Amazon Armbench dataset (https://armbench.s3.amazonaws.com/index.html).
TotalSegmentator-MR-Lite
About
This is a derivative of the TotalSegmentator dataset
616 MR images and corresponding segmentation mask of 50 structures
We combined multiple segmentation masks into a single nii.gz file under the folder Masks,
and moved all MR images to the folder Images.
All images and masks are renamed according to case IDs.
This dataset is released under the CC BY-NC-SA 2.0 license.
News 🔥
[10 Oct, 2025] This dataset is integrated into 🔥MedVision🔥… See the full description on the dataset page: https://huggingface.co/datasets/YongchengYAO/TotalSegmentator-MR-Lite.thinking_toto_lerobot_output_qwen3vlarmbench-segmentation-mix-object-toteThis is data from the Amazon Armbench dataset (https://armbench.s3.amazonaws.com/index.html).
totaldrama-qwen-image-lora-datasetTotalDrmkpo-figurestotal_text_train_cleaned
total_text_train_cleaned
The total_text_train family of the ElliotVL supervised-fine-tuning pool, after VLM cleaning.
images
1,255
QA turns
5,031
answers rewritten by the cleaning pass
499
QA created by the cleaning pass (new_qa)
4,326 (86.0%)
shards
1
How this was cleaned
A vision-language model read each image together with its QA and judged the item. The pass is
not a filter that only removes rows — it rewrites answers it finds wrong… See the full description on the dataset page: https://huggingface.co/datasets/Elliot-Data/total_text_train_cleaned.pubmed_totalTotalText_RS_nothink
TotalText — TotalText_RS_nothink
Rejection-sampled from the TotalText train split. This split holds the accepted items, answer only.
rows
291
QA pairs
291
shards
4
accepted / rejected (whole family)
291 / 313
accept rate
48.2%
verifier
lines
How the data was produced
A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with
the official ground truth by the verifier described below; matches go to… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/TotalText_RS_nothink.Suim_totalconcrete_crack_combinedloloTotalText_RS_think
TotalText — TotalText_RS_think
Rejection-sampled from the TotalText train split. This split holds the accepted items, with the model's reasoning trace.
rows
291
QA pairs
291
shards
4
accepted / rejected (whole family)
291 / 313
accept rate
48.2%
verifier
lines
How the data was produced
A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with
the official ground truth by the verifier described… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/TotalText_RS_think.TotalText_rejected
TotalText — TotalText_rejected
Rejection-sampled from the TotalText train split. This split holds the rejected items — the answer field holds the official ground truth.
rows
313
QA pairs
313
shards
3
accepted / rejected (whole family)
291 / 313
accept rate
48.2%
verifier
lines
The rejected split is training data, not just diagnostics: answer is the official ground truth, and wrong_vlm records what the model said instead.
How the data was… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/TotalText_rejected.TotalText_longform
TotalText — TotalText_longform
Rejection-sampled from the TotalText train split. This split holds the long-form items the verifier cannot score, shipped unverified.
rows
61
QA pairs
61
shards
1
accepted / rejected (whole family)
291 / 313
accept rate
48.2%
verifier
lines
How the data was produced
A VLM answers every question at temperature 0 with reasoning enabled. Its answer is compared with
the official ground truth by the verifier… See the full description on the dataset page: https://huggingface.co/datasets/elliot-mllm/TotalText_longform.total-texttotaltextimnet1k_totem_poleimnet1k_redshank_Tringa_totanustotal_synthesis_of_new_drugs
Total Synthesis of New Drugs
This dataset contains basic information and raw synthesis route data for 60 drugs, along with 600 multimodal reasoning supervision (CoT) fine-tuning data.
The experimental and computational work in this dataset run on the Huawei Cloud AI Compute Service. We appreciate the stable compute supply from this platform.
新药化学全合成路线
该数据集包含 60 种药物的基本信息与合成路线原始数据,以及 600 条多模态思维监督微调数据。
本数据集的实验与计算工作依托于华为昇腾AI云服务平台完成,特此对其提供的稳定算力支持表示感谢。
triatominos-augmentado-parquetfixed_bin_fixed_tote_shirt_cassidy_bertha_yondu_10142025gym_xarm-XarmLift_v0-TOTAL_VLA
