CoolFace
26 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01milan477 /MuSP-Bench MuSP-Bench MuSP-Bench is a 490-question benchmark for musical score understanding, performance listening, and combined score-performance reasoning. Contents data/questions.csv: all 490 questions, accepted answers, and the response contract for each. inputs/pdf/without_context/: one context-removed PDF per piece. inputs/images/: rendered score-page images for every piece. inputs/abc/: one ABC score per piece. inputs/abc_plus_midi/: one aligned ABC+MIDI… See the full description on the dataset page: https://huggingface.co/datasets/milan477/MuSP-Bench.imagequestion-answeringn<1K1 likes190 downloads1mo agoHugging Face02milanow /PersonaMem-v2 PersonaMem-v2: Towards Personalized Intelligence via Learning Implicit User Personas and Agentic Memory 🚨 The paper is now released. View the full paper here and codebase here. 🙌 The dataset has been downloaded over 12,000 times. Thank you everybody for finding our work helpful! Personalization is becoming the next milestone of artificial super-intelligence. AI cannot always satisfy every user, especially on tasks with subjective goals, but personalization… See the full description on the dataset page: https://huggingface.co/datasets/milanow/PersonaMem-v2.tabularquestion-answering10K<n<100K0 likes102 downloads5mo agoHugging Face03mila-intel /ProtST-EnzymeCommissiontext10K<n<100K0 likes43 downloads2y agoHugging Face04mila-intel /ProtST-BinaryLocalizationtabular1K<n<10K1 likes38 downloads3y agoHugging Face05mila-intel /ProtDescribetext100K<n<1M0 likes31 downloads2y agoHugging Face06milamarcheva /bulgarian_cds_lg Dataset Overview A sentence-level corpus drawn from scanned Bulgarian children's text. Each row represents one segmented sentence, its tokenization, the source URL, and its token count. Data Schema Column Type Description MainSentencised string The raw, sentence-segmented text (in Bulgarian). TokenisedSent list[string] The sentence split into word-tokens (lowercased, stripped). SourceLink string (URL) Origin of the sentence (e.g. a Chitanka book/text ZIP).… See the full description on the dataset page: https://huggingface.co/datasets/milamarcheva/bulgarian_cds_lg.tabular1M<n<10M2 likes28 downloads1y agoHugging Face07mila-intel /ProtST-GeneOntology-CCtext10K<n<100K0 likes26 downloads2y agoHugging Face08MilaNLProc /survey-language-technologies The AI Gap: How Socioeconomic Status Affects Language Technology Interactions 🏆 Best Social Impact Paper Award at ACL 2025 Dataset Summary This dataset comprises responses from 1,000 individuals from diverse socioeconomic backgrounds, collected to study how socioeconomic status (SES) influences interaction with language technologies, particularly generative AI and large language models (LLMs). Participants shared demographic and socioeconomic data, as well as up to 10… See the full description on the dataset page: https://huggingface.co/datasets/MilaNLProc/survey-language-technologies.text1K<n<10K3 likes24 downloads1y agoHugging Face09mila-intel /ProtST-AAVtabular10K<n<100K0 likes23 downloads2y agoHugging Face10mila-intel /ProtST-Thermostabilitytabular1K<n<10K0 likes23 downloads2y agoHugging Face11mila-intel /ProtST-Fluorescencetabular10K<n<100K0 likes22 downloads2y agoHugging Face12mila-intel /ProtST-BetaLactamasetabular1K<n<10K0 likes20 downloads2y agoHugging Face13mila-intel /subloc_templatetextn<1K0 likes19 downloads3y agoHugging Face14mila-intel /ProtST-Stabilitytabular10K<n<100K0 likes17 downloads2y agoHugging Face15mila-intel /ProtST-GeneOntology-MFtext10K<n<100K0 likes15 downloads2y agoHugging Face16Milana /russian_keywordstextsummarization10K<n<100K1 likes12 downloads3y agoHugging Face17mila-intel /ProtST-GeneOntology-BPtext10K<n<100K0 likes11 downloads2y agoHugging Face18Milana /russian-indi-alternativetext1K<n<10K0 likes8 downloads3y agoHugging Face19milan44 /fashion-custom-datatext1K<n<10K0 likes6 downloads2y agoHugging Face20milan44 /fashion-updatedtext1K<n<10K0 likes5 downloads2y agoHugging Face21MilanSalvi /autotrain-xxhsn-fsq88textn<1K0 likes4 downloads2y agoHugging Face22milan44 /fashiontext1K<n<10K0 likes3 downloads2y agoHugging Face23milandev /mma_trainingtabularn<1K0 likes3 downloads4mo agoHugging Face24mila-ai4h /biasly-datagated Biasly: An Expert-Annotated Dataset for Subtle Misogyny Detection and Mitigation This repository contains the dataset presented in the paper, Biasly: An Expert-Annotated Dataset for Subtle Misogyny Detection and Mitigation. The dataset is the first of its kind in that it provides detailed annotations for each misogynsitic instance, including sub-categories of misogyny, a continuous severity score, and potentially a rewritten version of the original datapoit with the misogyny reduced… See the full description on the dataset page: https://huggingface.co/datasets/mila-ai4h/biasly-data.tabular1M<n<10M6 likes2 downloads2y agoHugging Face25milab12 /face2profileimage1K<n<10K0 likes2 downloads1y agoHugging Face26miladj3 /bibletextn<1K0 likes2 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.