CoolFace
Datasetpublicgated

BIOMEDICA/biomedica_webdataset_24M

Dataset Card for Dataset Name Arxiv: Arxiv     |     Website: Biomedica     |     Training instructions: OpenCLIP     |     Tutorial: Google Colab BIOMEDICA Dataset is a large-scale, deep-learning-ready biomedical dataset containing over 24M imagecaption pairs and 30M image-references from 6M unique open-source articles. Each… See the full description on the dataset page: https://huggingface.co/datasets/BIOMEDICA/biomedica_webdataset_24M.

sourceHugging Faceupdated 26d agoView on Hugging Face
40likes4.2kdownloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
BIOMEDICA/biomedica_webdataset_24M · CoolFace