datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
SynthCheX-75K-v2
SynthCheX-75K
SynthCheX-75K is released as a part of the CheXGenBench paper. It is a synthetic dataset generated using Sana (0.6B) [1] fine-tuned on chest radiographs. Sana (0.6B) establishes the SoTA performance on the CheXGenBench benchmark.
The dataset contains 75,649 high-quality image-text samples along with the pathological annotations.
Filtration Process for SynthCheX-75K
Generative models can lead to both high and low-fidelity generations on different subsets… See the full description on the dataset page: https://huggingface.co/datasets/raman07/SynthCheX-75K-v2.Diabetic_Retinopathy_Preprocessed_Dataset_256x256This is dataset comes from this Kaggle Dataset
from the user Sachin Kumar.
The goal of the dataset is for the Varun AIM Projects to easily start running and download the dataset on their local computer in the HF libraries as the directory I strongly recommedn to use.
