datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
showdown-clicks
showdown-clicks
General Agents
🤗 Dataset | GitHub
showdown is a suite of offline and online benchmarks for computer-use agents.
showdown-clicks is a collection of 5,679 left clicks of humans performing various tasks in a macOS desktop environment. It is intended to evaluate instruction-following and low-level control capabilities of computer-use agents.
As of March 2025, we are releasing a subset of the full set, showdown-clicks-dev, containing 557 clicks. All examples are… See the full description on the dataset page: https://huggingface.co/datasets/generalagents/showdown-clicks.CLIP-ViT-L-14-336-L20-features
OpenAI/CLIP-ViT-L/14@336 Layer 20 features, CLIP+BLIP labels
Feature activation max visualization of the 4096 Features @ L20
CLIP+BLIP labels (may or may not describe what a neuron truly encodes!)
⚠️ May contain sensitive images, albeit abstract. Use responsibly!
Examples:
WIT-es_jina-clip-v2_sampleclinical-narrative-image-integrity-v0.2
Clinical Narrative Image Integrity v0.2
What this is
A small dataset that tests one question:
Can you detect when a clinical narrative-image system is moving toward integrity failure, not just carrying ambiguity?
This repo focuses on narrative-image integrity under clinical reasoning pressure.
It models a system where:
narrative coherence may weaken
image alignment may drift
interpretive distortion may rise
fragmented signal may destabilize representation before overt… See the full description on the dataset page: https://huggingface.co/datasets/ClarusC64/clinical-narrative-image-integrity-v0.2.MMQSD_ClipSyntel
Dataset Card for MMCQS Dataset
This is the MMCQS Dataset that have been used in the paper "CLIPSyntel: CLIP and LLM Synergy for Multimodal Question Summarization in Healthcare"
Github: https://github.com/AkashGhosh/CLIPSyntel-AAAI2024
Paper: https://arxiv.org/pdf/2312.11541
Uses
Download and unzip the Multimodal_images_finalnew.zip file, that can be found the in the 'Files and Version' section, to access the images that have been used in the dataset. The image… See the full description on the dataset page: https://huggingface.co/datasets/ArkaAcharya/MMQSD_ClipSyntel.ccs_synthetic_translated_arabic_processedccs_synthetic_translated_arabicThe columns inside the dataset as follows:
index
url
caption_en
caption_ar
The dataset size is 12556500 rows × 4 columns
clip_finetuning_dataset
