CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ZachSun /video-lvlm-datatext10K<n<100K0 likes195 downloads2y agoHugging Face02EricY05 /lvlm-ood-fake-dataimage10K<n<100K0 likes160 downloads3mo agoHugging Face03Intel /Uncovering_LVLM_Biastabular10M<n<100M0 likes155 downloads1y agoHugging Face04YangyiYY /LVLM_NLFNOTE: LVLM_NLF and VLSafe are constructed based on COCO and LLaVA. So the image can be directly retrieved from the COCO train-2017 version using the image id. LVLM_NLF (Large Vision Language Model with Natural Language Feedback) Dataset Card Dataset details Dataset type: LVLM_NLF is a GPT-4-Annotated natural language feedback dataset that aims to improve the 3H alignment and interaction ability of large vision-language models (LVLMs). Dataset date: LVLM_NLF was collected between September… See the full description on the dataset page: https://huggingface.co/datasets/YangyiYY/LVLM_NLF.text-generation10K<n<100K12 likes89 downloads3y agoHugging Face05GenIntelLab /SOCO-LVLM SOCO-LVLM SOCO-LVLM provides multiple-choice semantic object correspondence evaluation data for LVLMs. This is the SOCO-LVLM v1 release, derived from SOCOv1. The original SOCO correspondence benchmark is available in the GenIntelLab/SOCO dataset repository. Repository Layout GenIntelLab/SOCO-LVLM SOCO_LVLM/ soco_lvlm_img.tsv soco_lvlm_imgtxt.tsv soco_lvlm_txt.tsv README.md Variants soco_lvlm_img.tsv: image-input evaluation variant… See the full description on the dataset page: https://huggingface.co/datasets/GenIntelLab/SOCO-LVLM.visual-question-answering1 likes83 downloads1mo agoHugging Face06swap-uniba /Extending-LVLMs-NonEnglish-DataResource associated to the paper "Extending Large Language Models to Multimodality for non-English Languages". We provide the following resources: align/: projector alignment dataset of LLaVA translated to Italian and Spanish using madlad test/: contains the MultiInstruct test set formatted in English, as well as the formatted and translated version in Italian and Spanish using MadLad and human written formattings test_conversation/: contains the ImageDialog type 1 subset of the OmniDialog… See the full description on the dataset page: https://huggingface.co/datasets/swap-uniba/Extending-LVLMs-NonEnglish-Data.0 likes62 downloads9mo agoHugging Face07STIC-LVLM /stic-coco-preference-6kimage1K<n<10K0 likes24 downloads2y agoHugging Face08STIC-LVLM /stic-llava-instruct-desc-5ktext1K<n<10K0 likes22 downloads2y agoHugging Face09xiaoying0505 /LVLM_InterpretationThis repository contains the IDs of a subset of question used in the project: Where do Large Vision-Language Models Look at when Answering Questions? [paper] [code] It is a heatmap visualization method for interpreting Large Vision-Language Models (LVLMs) when generating open-ended answers. The original datasets can be obtained at CV-Bench, MMVP, MMStar. We sincerely appreciate the authors of these datasets for their contributions. This selected subseted is based on the relevance of the… See the full description on the dataset page: https://huggingface.co/datasets/xiaoying0505/LVLM_Interpretation.1K<n<10K0 likes16 downloads2y agoHugging Face10lvlm-anomaly-detection /datasets SEM anomaly-detection datasets Datasets for reproducing the SEM anomaly-detection results (detection + PALM domain adaptation). Delivered as *.zip files; download and unzip into data/datasets/. Contents file size what it is used for miic_test.zip 14G MIIC SEM test set — 5,000 normal + 116 abnormal MIIC detection eval oasis.zip 3.9G OASIS instruction-tuning data — synthetic particle/cut/bridge defects on normal SEM (per-class *_llava.json + images)… See the full description on the dataset page: https://huggingface.co/datasets/lvlm-anomaly-detection/datasets.image10K<n<100K0 likes6 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.