CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01manojkumarainala152 /ASVspoof2021_DF ASVspoof 2021 DF Benchmark-ready packaging of the DeepFake (DF) evaluation partition from ASVspoof 2021 for speech anti-spoofing and synthetic / deepfake voice detection. Overview This dataset contains the DF evaluation subset of the ASVspoof 2021 challenge. The task is binary classification: bonafide (genuine human speech) vs. spoof (synthetic, converted, or otherwise manipulated speech). The original dataset is available at… See the full description on the dataset page: https://huggingface.co/datasets/manojkumarainala152/ASVspoof2021_DF.audioaudio-classification100K<n<1M0 likes257 downloads16d agoHugging Face02manojkumarcs /indic-diarbench Indic DiarBench A multilingual joint diarization and ASR benchmark for Indian languages, spanning all 22 scheduled languages of India with approximately 108 hours of natural multi-speaker audio. Dataset Summary Indic DiarBench is a conversational speech benchmark designed to evaluate speaker-attributed ASR in realistic multi-speaker settings for Indian languages. All annotations are human-corrected with time-aligned, speaker-attributed transcriptions. The dataset… See the full description on the dataset page: https://huggingface.co/datasets/manojkumarcs/indic-diarbench.audioautomatic-speech-recognition1K<n<10K0 likes148 downloads2mo agoHugging Face03manoj8890 /big_patent Dataset Card for Big Patent Dataset Summary BIGPATENT, consisting of 1.3 million records of U.S. patent documents along with human written abstractive summaries. Each US patent application is filed under a Cooperative Patent Classification (CPC) code. There are nine such classification categories: a: Human Necessities b: Performing Operations; Transporting c: Chemistry; Metallurgy d: Textiles; Paper e: Fixed Constructions f: Mechanical Engineering; Lightning; Heating;… See the full description on the dataset page: https://huggingface.co/datasets/manoj8890/big_patent.textsummarization1M<n<10M0 likes139 downloads6mo agoHugging Face04manojkommineni /ca-sco-properties CA SCO Unclaimed Property California State Controller's Office unclaimed property records. Bucketed into alphabetical splits by owner last name first letter so every split stays under HF's 5 GB filter-index limit. Updated weekly via GitHub Actions. text10M<n<100M0 likes135 downloads7d agoHugging Face05KomeijiForce /Manosaba_Benchmark_Reorderedtext10K<n<100K0 likes107 downloads1mo agoHugging Face06manot /football-players Dataset Labels ['football', 'player'] Number of Images {'valid': 87, 'train': 119} How to Use Install datasets: pip install datasets Load the dataset: from datasets import load_dataset ds = load_dataset("manot/football-players", name="full") example = ds['train'][0] Roboflow Dataset Page https://universe.roboflow.com/konstantin-sargsyan-wucpb/football-players-2l81z/dataset/1 Citation @misc{… See the full description on the dataset page: https://huggingface.co/datasets/manot/football-players.imageobject-detectionn<1K1 likes95 downloads3y agoHugging Face07manot /pothole-segmentation Dataset Labels ['potholes', 'object', 'pothole', 'potholes'] Number of Images {'valid': 157, 'test': 80, 'train': 582} How to Use Install datasets: pip install datasets Load the dataset: from datasets import load_dataset ds = load_dataset("manot/pothole-segmentation", name="full") example = ds['train'][0] Roboflow Dataset Page https://universe.roboflow.com/abdulmohsen-fahad-f7pdw/road-damage-xvt2d/dataset/3 Citation… See the full description on the dataset page: https://huggingface.co/datasets/manot/pothole-segmentation.imageobject-detectionn<1K1 likes82 downloads3y agoHugging Face08Manoharareddy /MedHallu Dataset Card for MedHallu MedHallu is a comprehensive benchmark dataset designed to evaluate the ability of large language models to detect hallucinations in medical question-answering tasks. Dataset Details Dataset Description MedHallu is intended to assess the reliability of large language models in a critical domain—medical question-answering—by measuring their capacity to detect hallucinated outputs. The dataset includes two distinct splits:… See the full description on the dataset page: https://huggingface.co/datasets/Manoharareddy/MedHallu.text10K<n<100K0 likes73 downloads21d agoHugging Face09yuzhench /gigahands-vitra-mano GigaHands → VITRA Stage-1, official-MANO annotations VITRA Stage-1 hand annotations for GigaHands, with all joint positions taken from GigaHands' official MANO fit instead of mixing in triangulated keypoints. Annotations only — no videos (get those from GigaHands; the mapping is described in §5). episodes 13,247 (train 11,904 / test 1,343) frames 3,395,733 camera brics-odroid-001_cam0 (static rig; one constant extrinsic per scene) source GigaHands params/ +… See the full description on the dataset page: https://huggingface.co/datasets/yuzhench/gigahands-vitra-mano.tabularroboticsn<1K0 likes63 downloads2mo agoHugging Face10KomeijiForce /Manosaba_Benchmarktext10K<n<100K0 likes59 downloads2mo agoHugging Face11manojdahal191gom /claude-opus-4.6-4.7-reasoning-8.7k Background Ended up with some tokens to burn on a Claude Max plan. Assembly began during 4.6 and moved to 4.7. Model is tagged. The development evolved as it went along. The dataset has not been manually reviewed. It's entirely Claude developed. Clarification on Reasoning The reasoning is not Claude's actual chain-of-thought (cot) and is not summarized cot. It's a fully synthetic cot created as part of the Assistant response to mimic the type of "thinking"… See the full description on the dataset page: https://huggingface.co/datasets/manojdahal191gom/claude-opus-4.6-4.7-reasoning-8.7k.texttext-generation10K<n<100K0 likes54 downloads4mo agoHugging Face12manoelalmeida-io /github-pullrequeststabularn<1K1 likes51 downloads2y agoHugging Face13manoela /noticias_ptbrtext100K<n<1M0 likes46 downloads1y agoHugging Face14manot /pothole-segmentation2 Dataset Labels ['pothole'] Number of Images {'valid': 133, 'test': 66, 'train': 466} How to Use Install datasets: pip install datasets Load the dataset: from datasets import load_dataset ds = load_dataset("manot/pothole-segmentation2", name="full") example = ds['train'][0] Roboflow Dataset Page https://universe.roboflow.com/gurgen-hovsepyan-mbrnv/pothole-detection-gilij/dataset/2 Citation @misc{… See the full description on the dataset page: https://huggingface.co/datasets/manot/pothole-segmentation2.imageobject-detectionn<1K1 likes44 downloads3y agoHugging Face15fracapuano /manus-mano-poses Manus MANO Poses This dataset contains a right-hand Manus glove recording converted into the 21-landmark hand-pose convention used by the orca_teleop pipeline and its retargeters. Contents Split: train Frames: 2278 Duration: 37.95 s Sampling rate: 60.00 Hz Handedness: right Each row contains frame, elapsed timestamps, handedness, and keypoints, a (21, 3) float32 array of wrist-relative 3D landmarks in meters. Landmark Order The keypoints array follows the… See the full description on the dataset page: https://huggingface.co/datasets/fracapuano/manus-mano-poses.tabularrobotics1K<n<10K1 likes44 downloads4mo agoHugging Face16ManojKhanal /control_image2image10K<n<100K0 likes42 downloads2y agoHugging Face17manojdec25 /diamond-price-predictor-logs2 Dataset Card for Dataset Name Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed]… See the full description on the dataset page: https://huggingface.co/datasets/manojdec25/diamond-price-predictor-logs2.tabularn<1K0 likes41 downloads2y agoHugging Face18manoj-dhakal /philosloppy_encyclopediatext100K<n<1M3 likes39 downloads2y agoHugging Face19manojbalaji1 /anveshana Dataset Card for Anveshana Dataset Details Dataset Description we embarked on a comprehensive benchmarking study to explore and evaluate current state-of-the-art models for Cross-Lingual Information Retrieval (CLIR) from English to Sanskrit. Our primary objective is to assess the effectiveness of these models in accurately retrieving Sanskrit documents based on English queries. To achieve this, we meticulously assembled a robust dataset, focusing on the… See the full description on the dataset page: https://huggingface.co/datasets/manojbalaji1/anveshana.texttext-retrieval10K<n<100K1 likes39 downloads1y agoHugging Face20bochen123 /ego4d-hand-manogated Dataset summary Per-frame 3D hand annotations and text captions for 1,302,538 egocentric video clips drawn from 3,320 Ego4D videos. Every clip is 121 frames at 30 fps (4.03 s) at a 540-pixel short side. This is the annotation release accompanying Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints (ECCV 2026). Each clip carries, for both hands and every frame: 21 3D joints, MANO pose parameters, root rotation and translation, 2D wrist position and… See the full description on the dataset page: https://huggingface.co/datasets/bochen123/ego4d-hand-mano.tabulartext-to-video1M<n<10M1 likes38 downloads16d agoHugging Face21manojkumarvohra /replicated_emotions Dataset Summary Emotion is a dataset of English Twitter messages with six basic emotions: anger, fear, joy, love, sadness, and surprise. For more detailed information please refer to the paper. This dataset is a processed form of "dair-ai/emotion" dataset. [https://huggingface.co/datasets/dair-ai/emotion] In this one, I have replicated/duplicated the samples for minority classes so that all the emotion classes have [approximate] equal sample count. dataset_info: features: name:… See the full description on the dataset page: https://huggingface.co/datasets/manojkumarvohra/replicated_emotions.text10K<n<100K0 likes37 downloads3y agoHugging Face22open-llm-leaderboard /ManoloPueblo__LLM_MERGE_CC2-detailsgated Dataset Card for Evaluation run of ManoloPueblo/LLM_MERGE_CC2 Dataset automatically created during the evaluation run of model ManoloPueblo/LLM_MERGE_CC2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ManoloPueblo__LLM_MERGE_CC2-details.tabular10K<n<100K0 likes36 downloads2y agoHugging Face23manoela /plantvillageimage10K<n<100K0 likes35 downloads2y agoHugging Face24open-llm-leaderboard /ManoloPueblo__LLM_MERGE_CC3-detailsgated Dataset Card for Evaluation run of ManoloPueblo/LLM_MERGE_CC3 Dataset automatically created during the evaluation run of model ManoloPueblo/LLM_MERGE_CC3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ManoloPueblo__LLM_MERGE_CC3-details.tabular10K<n<100K0 likes34 downloads2y agoHugging Face25open-llm-leaderboard /ManoloPueblo__ContentCuisine_1-7B-slerp-detailsgated Dataset Card for Evaluation run of ManoloPueblo/ContentCuisine_1-7B-slerp Dataset automatically created during the evaluation run of model ManoloPueblo/ContentCuisine_1-7B-slerp The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ManoloPueblo__ContentCuisine_1-7B-slerp-details.tabular10K<n<100K0 likes32 downloads2y agoHugging Face26manojkumarvohra /amplified_emotions Dataset Summary Emotion is a dataset of English Twitter messages with six basic emotions: anger, fear, joy, love, sadness, and surprise. For more detailed information please refer to the paper. This dataset is a processed form of "dair-ai/emotion" dataset. [https://huggingface.co/datasets/dair-ai/emotion] In this one, I have amplified the samples for minority classes so that all the emotion classes have [approximate] equal sample count. There is another dataset with duplicate… See the full description on the dataset page: https://huggingface.co/datasets/manojkumarvohra/amplified_emotions.text10K<n<100K1 likes29 downloads3y agoHugging Face27Manoj2702 /Sanskrit-to-English-Vocabulary-v1 Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: Manoj, Nandish, Mayank, Abhiram Language(s) (NLP): Sanskrit, English License: MIT License Repository: [More Information Needed] Paper [optional]: [More Information Needed] Demo [optional]: [More Information Needed] Uses Direct Use [More… See the full description on the dataset page: https://huggingface.co/datasets/Manoj2702/Sanskrit-to-English-Vocabulary-v1.text100K<n<1M1 likes27 downloads2y agoHugging Face28KomeijiForce /Manosaba_Scriptstabular10K<n<100K0 likes27 downloads2mo agoHugging Face29mano-wii /blender_duplicates Dataset Card for Dataset Name Contains reduced description of issues reported at https://projects.blender.org/blender/blender/issues and points to duplicate issues in order to categorize similarity. This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Each report has been shortened by removing frequently repeated texts such as System Information, Blender Version… See the full description on the dataset page: https://huggingface.co/datasets/mano-wii/blender_duplicates.texttext-classification1K<n<10K0 likes26 downloads3y agoHugging Face30Manoj0002manoj /salesdatatabularn<1K0 likes26 downloads25d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.