datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
tripmatch-ai-dataset
TripMatch AI Dataset
A reproducible multimodal dataset for the TripMatch AI Final Project. It contains
10,000 synthetic text trip plans with a raw idea generated for every row by the
pretrained Hugging Face model google/flan-t5-small, plus 5,000 real street-view images
retained as extra multimodal work. The two configurations are separate so Dataset
Viewer can load each schema correctly.
Dataset statistics
Configuration
Rows
Main fields
Intended task… See the full description on the dataset page: https://huggingface.co/datasets/avihayamor/tripmatch-ai-dataset.cxr14-bpals-trial
NIH-CXR14-BPALS — Label-Quality Audit for NIH ChestX-ray14
By O5I | Patent pending (B-PALS) | Contact: hello@o5i.io
Start by finding what's wrong. Refine from there.
What it does
NIH ChestX-ray14's labels are derived automatically from radiology reports, not verified against the images — so a meaningful share are noisy or wrong (~80% reported accuracy). NIH-CXR14-BPALS independently re-examines each (image, label) pair with a vision-language model and returns a… See the full description on the dataset page: https://huggingface.co/datasets/o5i/cxr14-bpals-trial.TriALS-Report
TriALS-Report: A Multi-Center Benchmark for Abdominal Disease Diagnosis and Report Generation from Non-Contrast CT
Study workflow. Non-contrast CT volumes are paired with the triphasic contrast-enhanced report of the same patient; findings are extracted from the report to form the label space, and models are evaluated on disease diagnosis and report generation.
TriALS-Report is a multi-centre benchmark for abdominal disease diagnosis from non-contrast CT (NCCT), where the… See the full description on the dataset page: https://huggingface.co/datasets/marwankefah/TriALS-Report.
