CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01wovenbytoyota-vai /CaST-Bench CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering This is the official repository for the CaST-Bench dataset, introduced in the paper "CaST-Bench: Benchmarking Causal Chain-Grounded Spatio-Temporal Reasoning for Video Question Answering". CaST-Bench is the first benchmark to evaluate Vision-Language Models (VLMs) on causal chain reasoning grounded in fine-grained spatio-temporal evidence. Given a video and a causal question… See the full description on the dataset page: https://huggingface.co/datasets/wovenbytoyota-vai/CaST-Bench.textvideo-text-to-textn<1K0 likes450 downloads3mo agoHugging Face02wovenbytoyota-vai /InstVL InstVL: A Large-Scale Instance-Aware Vision-Language Dataset This is the official repository for the InstVL dataset, introduced in the paper InstAP: Instance-Aware Vision-Language Pre-Train for Spatial-Temporal Understanding. InstVL is a large-scale dataset of images and videos designed to bridge the gap between holistic scene understanding and fine-grained, instance-level comprehension. Current vision-language pre-training (VLP) paradigms excel at global scene understanding but… See the full description on the dataset page: https://huggingface.co/datasets/wovenbytoyota-vai/InstVL.textimage-to-text1M<n<10M5 likes284 downloads6mo agoHugging Face03vaidehi99 /UnLOK-VQA 📊 Dataset: UnLOK-VQA (Unlearning Outside Knowledge VQA) Paper: Unlearning Sensitive Information in Multimodal LLMs: Benchmark and Attack-Defense Evaluation Code: https://github.com/Vaidehi99/mmmedit Link: Dataset Link This dataset contains approximately 500 entries with the following key attributes: "id": Unique Identifier for each entry "src": The question whose answer is to be deleted ❓ "pred": The answer to the question meant for deletion ❌ "loc": Related neighborhood questions… See the full description on the dataset page: https://huggingface.co/datasets/vaidehi99/UnLOK-VQA.tabularvisual-question-answeringn<1K0 likes54 downloads1y agoHugging Face04vaibhav2507 /hdfs-logstexttext-classification100K<n<1M0 likes52 downloads1y agoHugging Face05vaibhav2507 /bgl-logstexttext-classification1M<n<10M0 likes43 downloads1y agoHugging Face06vai-org /EC-Benchtabular1K<n<10K0 likes32 downloads6mo agoHugging Face07VAIBHAV22334455 /NOVA50ktext10K<n<100K0 likes31 downloads2y agoHugging Face08Vaibhav-GOAT /nepi-prompts-dataset NEPI: Narrative-Embedded Prompt Injection Dataset (Sanitized) Dataset Summary This dataset contains 4,000 sanitized prompts designed for research on prompt injection vulnerabilities in Large Language Models (LLMs).It introduces and supports evaluation of a novel attack class called Narrative-Embedded Prompt Injection (NEPI), where adversarial intent is embedded inside coherent fictional narratives, dialogues, or persona-driven roleplay prompts. Unlike traditional… See the full description on the dataset page: https://huggingface.co/datasets/Vaibhav-GOAT/nepi-prompts-dataset.texttext-generation1K<n<10K0 likes23 downloads8mo agoHugging Face09VaisakhKrishna /Emotional_Sentiment_AnalysisEmotional Sentiment Analysis Dataset for LLaMA-2 Fine-tuning (The formatted version can be directly used for fine tuning which contain only the formatted text, while the dataset.csv contain all the text, emotion, response and the formatted text) This dataset contains conversational data for training and fine-tuning language models for emotional sentiment analysis and response generation. The dataset includes user inputs, their corresponding emotional states, and tailored chatbot responses… See the full description on the dataset page: https://huggingface.co/datasets/VaisakhKrishna/Emotional_Sentiment_Analysis.texttext-classification1K<n<10K2 likes15 downloads2y agoHugging Face10jadhavmanasi70 /adaption-vaidya-rural-symptoms This dataset is a remastered version prepared using Adaption's Adaptive Data platform. adaption-vaidya_rural_symptoms This dataset maps colloquial symptom expressions from multiple Indian languages and dialects to standardized medical meanings and severity levels. It is designed to bridge the gap between rural healthcare communication and formal medical terminology for NLP applications. The data includes core fields for symptom phrases, language, dialect, corrected meaning, and… See the full description on the dataset page: https://huggingface.co/datasets/jadhavmanasi70/adaption-vaidya-rural-symptoms.text1K<n<10K0 likes15 downloads4mo agoHugging Face11VAIBHAV22334455 /JARVIStextn<1K0 likes14 downloads2y agoHugging Face12jadhavmanasi70 /adaption-vaidya-rural-symptoms-v1 This dataset is a remastered version prepared using Adaption's Adaptive Data platform. adaption-vaidya_rural_symptoms This dataset maps colloquial symptom expressions from multiple Indian languages and dialects to standardized medical meanings and severity levels. It is designed to bridge the gap between rural healthcare communication and formal medical terminology for NLP applications. The data includes core fields for symptom phrases, language, dialect, corrected meaning, and… See the full description on the dataset page: https://huggingface.co/datasets/jadhavmanasi70/adaption-vaidya-rural-symptoms-v1.text10K<n<100K0 likes13 downloads3mo agoHugging Face13vaidehi99 /unltabular1K<n<10K0 likes12 downloads2y agoHugging Face14vaishnavipadmanabhan /adaption-chest-xray-pneumonia-labels This dataset is a remastered version of this dataset prepared using Adaption's Adaptive Data platform. adaption-chest_xray_pneumonia_labels This dataset contains labeled samples for chest X-ray image classification, distinguishing between normal cases and those with pneumonia. Each entry provides a diagnostic label indicating the presence of pneumonia or a normal lung condition. The data is structured as pairs with a single completion field holding the categorical diagnosis.… See the full description on the dataset page: https://huggingface.co/datasets/vaishnavipadmanabhan/adaption-chest-xray-pneumonia-labels.image1K<n<10K1 likes12 downloads3mo agoHugging Face15vaibhav-mugulavalli-2004 /ToolCall-SFT-v2-40Ktabular10K<n<100K0 likes10 downloads1mo agoHugging Face16VAIBHAV22334455 /EMOTIONAL-JARVIStext1K<n<10K1 likes9 downloads2y agoHugging Face17vaidehi99 /HC_concepts_refusaltext1K<n<10K0 likes9 downloads2y agoHugging Face18vaibhavbasidoni /unsloth-sharagatextn<1K0 likes9 downloads1y agoHugging Face19vairontalvix /hindicorrectiontextn<1K0 likes7 downloads2y agoHugging Face20vaidehi99 /combined_refusal_datatext1K<n<10K0 likes6 downloads2y agoHugging Face21vaikhari-ai /coregated CORE: Comprehensive Ontological Relation Evaluation 🌐 Website | 📄 Paper | 💻 Code Dataset Summary CORE is a human-grounded benchmark for evaluating large language models on fundamental semantic and ontological reasoning. It assesses whether models can correctly recognize a broad range of sense-level relations and, critically, identify when no meaningful relationship exists between concepts. With comprehensive relation coverage and strong human baselines… See the full description on the dataset page: https://huggingface.co/datasets/vaikhari-ai/core.textmultiple-choicen<1K3 likes6 downloads8mo agoHugging Face22vaidehi99 /LC_concepts_refusaltext1K<n<10K0 likes5 downloads2y agoHugging Face23vaibhavbasidoni /dataset7textn<1K0 likes4 downloads1y agoHugging Face24vaibhavbasidoni /charak-chapter2textn<1K0 likes4 downloads1y agoHugging Face25vaibhavkesarwani /gitsolve_alpaca_datasettext10K<n<100K0 likes4 downloads1y agoHugging Face26vaidehi99 /MC_concepts_refusaltext1K<n<10K0 likes3 downloads2y agoHugging Face27vaishnavin /medical.jsontextn<1K1 likes3 downloads1y agoHugging Face28vaish026 /alpaca-cleaned-10k-chattext10K<n<100K0 likes3 downloads7mo agoHugging Face29Vaibhav9401 /testllamatextn<1K0 likes2 downloads3y agoHugging Face30Vaishali111 /story-datasettextn<1K0 likes2 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.