CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01DAMO-NLP-SG /multimodal_textbook Multimodal-Textbook-6.5M Overview This dataset is for "2.5 Years in Class: A Multimodal Textbook for Vision-Language Pretraining", containing 6.5M images interleaving with 0.8B text from instructional videos. It contains pre-training corpus using interleaved image-text format. Specifically, our multimodal-textbook includes 6.5M keyframesextracted from instructional videos, interleaving with 0.8B ASR texts. All the images and text are extracted from online… See the full description on the dataset page: https://huggingface.co/datasets/DAMO-NLP-SG/multimodal_textbook.text-generation1M<n<10M164 likes5.3k downloads2y agoHugging Face02sjidnd /C_damoxing0 likes4.2k downloads2y agoHugging Face03DAMO-NLP-SG /VideoRefer-700K VideoRefer-700K Paper | Project Page | Code VideoRefer-700K is a large-scale, high-quality object-level video instruction dataset. Curated using a sophisticated multi-agent data engine to fill the gap for high-quality object-level video instruction data. VideoRefer consists of three types of data: Object-level Detailed Caption Object-level Short Caption Object-level QA Video sources: Detailed&Short Caption Panda-70M. QA MeViS A2D Youtube-VOS Data format: [ {… See the full description on the dataset page: https://huggingface.co/datasets/DAMO-NLP-SG/VideoRefer-700K.video-text-to-text8 likes3.7k downloads11mo agoHugging Face04sjidnd /LIb_damoxing15 likes3.1k downloads2y agoHugging Face05Alibaba-DAMO-Academy /ClinFusion-Eval-Data 🏥 ClinFusion-Eval-Data The Holistic Evaluation Suite for Vision-Centric Medical Multimodal LLMs ClinFusion-Eval-Data is the unified evaluation corpus used to benchmark the ClinFusion model series (ClinFusion-8B, ClinFusion-32B). It packages 211,810 evaluation records spanning 22 public medical benchmarks into a single, consistently-formatted suite, together with 509 GiB of the underlying 2D images and native 3D CT volumes they refer to. The goal is reproducibility:… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-DAMO-Academy/ClinFusion-Eval-Data.visual-question-answering100K<n<1M3 likes2.5k downloads1mo agoHugging Face06sjidnd /LIb_damoxing21 likes1.2k downloads2y agoHugging Face07DAMO-NLP-SG /MultiJail Multilingual Jailbreak Challenges in Large Language Models This repo contains the data for our paper "Multilingual Jailbreak Challenges in Large Language Models". [Github repo] Annotation Statistics We collected a total of 315 English unsafe prompts and annotated them into nine non-English languages. The languages were categorized based on resource availability, as shown below: High-resource languages: Chinese (zh), Italian (it), Vietnamese (vi) Medium-resource languages:… See the full description on the dataset page: https://huggingface.co/datasets/DAMO-NLP-SG/MultiJail.textn<1K12 likes1k downloads3y agoHugging Face08Alibaba-DAMO-Academy /RynnBrain-Bench RynnBrain-Bench Introduction We introduce RynnBrain-Bench, a high-dimensional evaluation suite designed to holistically benchmark the cognition and localization capabilities of embodied understanding models in complex household environments. Advancing beyond existing benchmarks, RynnBrain-Bench features a unique emphasis on fine-grained understanding and precise spatiotemporal localization within episodic video sequences. RynnBrain-Bench systematically… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-DAMO-Academy/RynnBrain-Bench.textvisual-question-answering10K<n<100K14 likes958 downloads7mo agoHugging Face09DAMO-NLP-MT /multialpacatext100K<n<1M12 likes654 downloads3y agoHugging Face10damo-da /oag-nepal-audit-reports OAG Nepal Audit Reports — Nepali transcripts and ruled tables Machine-readable transcripts of 6,234 publications of the Office of the Auditor General of Nepal (महालेखा परीक्षकको कार्यालय, OAG) — the annual audit reports of local governments, provinces and central bodies, plus the OAG's own bulletins, journals and financial statements. The OAG publishes these as PDFs whose text layer is, for most documents, legacy pre-Unicode Devanagari: fonts like Preeti and Fontasy Himali that… See the full description on the dataset page: https://huggingface.co/datasets/damo-da/oag-nepal-audit-reports.documenttext-retrieval10M<n<100M0 likes578 downloads23d agoHugging Face11DAMO-NLP-SG /Qwen2.5-7B-LongPO-128K-tokenized10K<n<100K0 likes418 downloads2y agoHugging Face12DAMO-NLP-SG /Multi-Source-Video-Captioning Multi-source Video Captioning (MSVC) Dataset Card Dataset details Dataset type: MSVC is a set of collected video captioning data. It is constructed to ensure a robust and thorough evaluation of Video-LLMs' video-captioning capabilities. Dataset detail: MSVC is introduced to address limitations in existing video caption benchmarks, MSVC samples a total of 1,500 videos with human-annotated captions from MSVD, MSRVTT, and VATEX, ensuring diverse scenarios and domains.… See the full description on the dataset page: https://huggingface.co/datasets/DAMO-NLP-SG/Multi-Source-Video-Captioning.textvisual-question-answering1K<n<10K7 likes323 downloads2y agoHugging Face13damonsalvatore123 /LMA-Individual-Projectimagen<1K0 likes309 downloads9d agoHugging Face14Alibaba-DAMO-Academy /InterVBench Video Drift Evaluation (vde.py) This repository contains a single entry point, vde.py, that computes Video Drift Error (VDE) scores for every .mp4 file inside a target directory. VDE provides a simple way to monitor how quality-related metrics drift across chunks of the same video. The script already supports several metric backends (clarity, motion, aesthetic, dynamic, subject, background) via the vbench tooling. Environment Setup Install the project… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-DAMO-Academy/InterVBench.2 likes254 downloads4mo agoHugging Face15DAMO-NLP-SG /Mistral-7B-LongPO-256K-tokenized10K<n<100K0 likes224 downloads2y agoHugging Face16rbalazs /DamogranLeftCup_20260911_120558This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranLeftCup_20260911_120558.tabularrobotics1K<n<10K0 likes191 downloads4d agoHugging Face17Alibaba-DAMO-Academy /RynnEC-Bench RynnEC-Bench RynnEC-Bench evaluates fine-grained embodied understanding models from the perspectives of object cognition and spatial cognition in open-world scenario. The benchmark includes 507 video clips captured in real household scenarios. Model Overall Mean Object Properties Seg. DR Seg. SR Object Mean Ego. His. Ego. Pres. Ego. Fut. World Size World Dis. World PR Spatial Mean GPT-4o 28.3 41.1 --- --- 33.9 13.4 22.8 6.0 24.3 16.7 36.1 22.2 GPT-4.1 33.5… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-DAMO-Academy/RynnEC-Bench.7 likes190 downloads1y agoHugging Face18rbalazs /DamogranMiddleCup_20260911_124118This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranMiddleCup_20260911_124118.tabularrobotics10K<n<100K0 likes184 downloads14d agoHugging Face19rbalazs /DamogranRightCup_20260911_140547This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup_20260911_140547.tabularrobotics1K<n<10K0 likes179 downloads14d agoHugging Face20DamonDemon /SpanUQ-Benchmark SpanUQ Benchmark A span-level uncertainty estimation benchmark for large language model generation. Each example contains an LLM-generated response decomposed into spans (contiguous text segments expressing single verifiable assertions), with uncertainty labels derived from sampling-based consistency verification. Quick Start from datasets import load_dataset # Load a specific model configuration ds = load_dataset("DamonDemon/SpanUQ-Benchmark", "Qwen3-14B")… See the full description on the dataset page: https://huggingface.co/datasets/DamonDemon/SpanUQ-Benchmark.tabulartext-generation10K<n<100K0 likes175 downloads3mo agoHugging Face21rbalazs /DamogranRightCup2_20260911_142724This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup2_20260911_142724.tabularrobotics1K<n<10K0 likes163 downloads14d agoHugging Face22DAMO-NLP-SG /Mistral-7B-LongPO-128K-tokenized10K<n<100K0 likes161 downloads2y agoHugging Face23rbalazs /DamogranLeftCup_20260911_122012This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranLeftCup_20260911_122012.tabularrobotics1K<n<10K0 likes160 downloads14d agoHugging Face24rbalazs /DamogranRightCup_20260911_130855This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup_20260911_130855.tabularrobotics1K<n<10K0 likes158 downloads14d agoHugging Face25rbalazs /DamogranRightCup_20260911_140000This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup_20260911_140000.tabularrobotics1K<n<10K0 likes158 downloads14d agoHugging Face26rbalazs /DamogranRightCup_20260911_130618This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup_20260911_130618.tabularrobotics1K<n<10K0 likes157 downloads14d agoHugging Face27rbalazs /DamogranRightCup_20260911_140335This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup_20260911_140335.tabularrobotics1K<n<10K0 likes156 downloads14d agoHugging Face28rbalazs /DamogranMiddleCup_20260911_130256This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranMiddleCup_20260911_130256.tabularrobotics1K<n<10K0 likes154 downloads14d agoHugging Face29rbalazs /DamogranRightCup_20260911_142409This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "shape": [ 6 ], "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos"… See the full description on the dataset page: https://huggingface.co/datasets/rbalazs/DamogranRightCup_20260911_142409.tabularrobotics1K<n<10K0 likes149 downloads14d agoHugging Face30Alibaba-DAMO-Academy /ClinHallu CLINHALLU Benchmark CLINHALLU is a benchmark for diagnosing stage-wise hallucinations in medical MLLM reasoning. Paper: CLINHALLU: A Benchmark for Diagnosing Stage-Wise Hallucinations in Medical MLLM ReasoningGitHub: alibaba-damo-academy/ClinHallu Benchmark Results Accuracy and stage-wise hallucination rates on CLINHALLU. We report answer accuracy (Acc) and hallucination rates for visual recognition (H^V), knowledge recall (H^K), and reasoning integration (H^R).… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-DAMO-Academy/ClinHallu.text10K<n<100K3 likes146 downloads3mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.