CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AL-GR /Origin-Sequence-Data AL-GR/Origin-Sequence-Data: Raw User Behavior Sequences 📜 About the Dataset Each row in this dataset (Origin-Sequence-Data) represents a step in a user's journey, consisting of a sequence of previously interacted items (user_history) and the next item they interacted with (target_item). All item IDs have been anonymized into short, unique strings. This dataset is ideal for: 🧑‍🔬 Researchers who want to design their own data processing or prompting strategies for… See the full description on the dataset page: https://huggingface.co/datasets/AL-GR/Origin-Sequence-Data.texttext-generation100K<n<1M0 likes362 downloads11mo agoHugging Face02laallein /OrigamIM OrigamIM: An Ambiguous Dataset of Sentence Interpretations, Implicit Moral Judgments and Reader Impressions Introduction Please cite following papers when using the origamIM dataset (paper 1 and paper 2): Allein, Liesbeth, and Marie-Francine Moens. "OrigamIM: An Ambiguous Dataset of Sentence Interpretations, Implicit Moral Judgments and Reader Impressions." Proceedings of the 3rd Workshop on Perspectivist Approaches to NLP @LREC-COLING 2024 (2024). Allein… See the full description on the dataset page: https://huggingface.co/datasets/laallein/OrigamIM.texttext-generation1K<n<10K0 likes85 downloads16d agoHugging Face03orionai /en_wikipedia_001 Dataset Card for en_wikipedia_001 The en_wikipedia_001 dataset is a collection of crawled paragraph text from Wikipedia on the 28th of April, 2024. It contains high-quality text, stored in multiple documents, available to be used to finetune or train AI models based that the license is followed. Dataset Details The dataset was crawled using our web crawler on the 28th of April at an average of 1 page per second as to respect robots.txt rules. Strict licensing must be… See the full description on the dataset page: https://huggingface.co/datasets/orionai/en_wikipedia_001.textquestion-answeringn<1K2 likes34 downloads2y agoHugging Face04MakiAi /Orin-Character-JP-v1texttext-generationn<1K0 likes10 downloads1y agoHugging Face05Orib24 /Roomly-Student-Bios-Multimodal Roomly: Multimodal Roommate Matching Dataset 🎯 Problem Statement Finding a roommate is often reduced to dry filters like "budget" and "location". Roomly aims to revolutionize this by focusing on personality, lifestyle, and visual preferences. This dataset provides synthetic student profiles and their ideal room environments. 📊 Exploratory Data Analysis (EDA) 1. User Persona Distribution Our dataset contains a balanced mix of different student… See the full description on the dataset page: https://huggingface.co/datasets/Orib24/Roomly-Student-Bios-Multimodal.texttext-generationn<1K0 likes7 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.