CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01ada-datadruids /full_tmdb_movies_datasettabular1M<n<10M8 likes555 downloads2y agoHugging Face02dirtycomputer /douban_movie_reviewtabular10M<n<100M9 likes479 downloads4y agoHugging Face03HenryWaltson /TMDB-IMDB-Movies-Datasettabular100K<n<1M1 likes328 downloads8mo agoHugging Face04arize-ai /movie_reviews_with_context_drift Dataset Card for reviews_with_drift Dataset Description Dataset Summary This dataset was crafted to be used in our tutorial [Link to the tutorial when ready]. It consists on a large Movie Review Dataset mixed with some reviews from a Hotel Review Dataset. The training/validation set are purely obtained from the Movie Review Dataset while the production set is mixed. Some other features have been added (age, gender, context) as well as a made up timestamp… See the full description on the dataset page: https://huggingface.co/datasets/arize-ai/movie_reviews_with_context_drift.tabulartext-classification10K<n<100K1 likes304 downloads4y agoHugging Face05ZhankuiHe /reddit_movie_large_v1 Dataset Card for Reddit-Movie-large-V1 Dataset Summary This dataset contains the recommendation-related conversations in movie domain, only for research use in e.g., conversational recommendation, long-query retrieval tasks. This dataset is ranging from Jan. 2012 to Dec. 2022. Another smaller version dataset (from Jan. 2022 to Dec. 2022) can be found here. Dataset Processing We dump Reddit conversations from pushshift.io, converted them into raw text on Reddit… See the full description on the dataset page: https://huggingface.co/datasets/ZhankuiHe/reddit_movie_large_v1.tabular1M<n<10M0 likes227 downloads3y agoHugging Face06wykonos /moviestabular100K<n<1M20 likes164 downloads3y agoHugging Face07ZhankuiHe /reddit_movie_small_v1 Dataset Card for Reddit-Movie-small-V1 Dataset Summary This dataset contains the recommendation-related conversations in movie domain, only for research use in e.g., conversational recommendation, long-query retrieval tasks. This dataset is ranging from Jan. 2022 to Dec. 2022. Another larger version dataset (from Jan. 2012 to Dec. 2022) can be found here. Dataset Processing We dump Reddit conversations from pushshift.io, converted them into raw text on Reddit… See the full description on the dataset page: https://huggingface.co/datasets/ZhankuiHe/reddit_movie_small_v1.tabular100K<n<1M1 likes142 downloads3y agoHugging Face08RummageLabs /pixar_movies Pixar Movies Dataset A comprehensive dataset of Pixar movies, including details on their release dates, directors, cast, box office performance, and ratings. This dataset is gathered from official sources, including Pixar, Rotten Tomatoes, and IMDb. For more information, visit Pixar. How the Data is Compiled All information in this dataset has been collected from public sources, including official information from Pixar, Rotten Tomatoes, and IMDb. Cells are each… See the full description on the dataset page: https://huggingface.co/datasets/RummageLabs/pixar_movies.tabularn<1K0 likes139 downloads2y agoHugging Face09bloc4488 /TMDB-all-moviestabular10K<n<100K0 likes138 downloads2y agoHugging Face10MangoGoes /douban_movie_info该数据集为豆瓣电影信息维表。 更多信息请参考文章《数据获取:豆瓣电影信息爬取》。 image10K<n<100K5 likes120 downloads3y agoHugging Face11wwbrannon /ml-interview-examples-movielens-1mtabular1M<n<10M0 likes118 downloads9mo agoHugging Face12johnidouglas /tmdb_5000_movies.csvTMDB 5000 Movie Dataset Original source: https://www.kaggle.com/datasets/tmdb/tmdb-movie-metadata tabular1K<n<10K0 likes98 downloads2y agoHugging Face13ShubhamChoksi /IMDB_Moviestabular1K<n<10K10 likes82 downloads3y agoHugging Face14reczoo /MovielensLatest_x1 MovielensLatest_x1 The MovieLens dataset consists of users' tagging records on movies. The task is formulated as personalized tag recommendation with each tagging record (user_id, item_id, tag_id) as an data instance. The target value denotes whether the user has assigned a particular tag to the movie. We provide the reusable, processed dataset released by the BARS benchmark, which are randomly split into 7:2:1 as the training set, validation set, and test set, respectively.… See the full description on the dataset page: https://huggingface.co/datasets/reczoo/MovielensLatest_x1.tabular1M<n<10M2 likes75 downloads3y agoHugging Face15tracywong117 /spam-douban-movie-review Description The Spam Douban Movie Reviews Dataset is a collection of movie reviews scraped from Douban, a popular Chinese social networking platform for movie enthusiasts. This dataset consists of reviews that have been manually classified as either spam or genuine by human reviewers. It contains a total of 1,600 data. This dataset is created for our project Spam Movie Reviews Detection through Supervised Learning. tabulartext-classification1K<n<10K5 likes69 downloads3y agoHugging Face16YUEMING2 /douban_movie_info该数据集为豆瓣电影信息维表。 更多信息请参考文章《数据获取:豆瓣电影信息爬取》。 image10K<n<100K1 likes56 downloads9mo agoHugging Face17drossi /EDA_on_IMDB_Movies_Datasetimagefeature-extraction1K<n<10K4 likes54 downloads3y agoHugging Face18codealchemist01 /letterboxd-movies Letterboxd Movies Dataset Dataset Description A comprehensive dataset of movies scraped from Letterboxd, including genres, ratings, runtime, countries, and detailed movie characteristics. This dataset contains 16246 movies with 28 features each, scraped from Letterboxd. It's perfect for: 🎬 Movie recommendation systems 📊 Film industry analysis 🤖 Machine learning projects 📈 Rating prediction models 🔍 Movie discovery algorithms Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/codealchemist01/letterboxd-movies.tabulartext-classification10K<n<100K1 likes53 downloads11mo agoHugging Face19shenmin91 /dirty-movie-a53fa0 dirty-movie-a53fa0 Synthetic sensors test data: 35 rows in data.csv. All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations. Fields sample_id: random identifier for this generated sample. row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/shenmin91/dirty-movie-a53fa0.tabularn<1K0 likes52 downloads13d agoHugging Face20eka416 /moviestabular1M<n<10M2 likes45 downloads1y agoHugging Face21pathii /IMDb_Top_250_Moviestabularn<1K1 likes43 downloads1y agoHugging Face22ada-datadruids /TMDB_movie_dataset_reducedtabular1M<n<10M0 likes40 downloads2y agoHugging Face23andreapisa9 /BLUE-Steer-MovieLens100Ktabular100K<n<1M0 likes40 downloads1mo agoHugging Face24santideleon /Rec-Gaze-Click-Cursor-Eye-Tracking-Movie-Recommendation-Dataset-for-Carousel-Interfaces RecGaze Dataset This is the HuggingFace RecGaze dataset from the paper: 'RecGaze: The First Eye Tracking and User Interaction Dataset for Carousel Interfaces'. Link to open-acess paper: SIGIR 2025 Dataset Description The RecGaze dataset is the first comprehensive feedback dataset on carousels that includes eye tracking results, clicks, cursor movements, and selection explanations. The dataset comprises of interactions from 3 &nbsp;movie selection tasks with 40… See the full description on the dataset page: https://huggingface.co/datasets/santideleon/Rec-Gaze-Click-Cursor-Eye-Tracking-Movie-Recommendation-Dataset-for-Carousel-Interfaces.tabularvisual-document-retrievaln<1K0 likes40 downloads26d agoHugging Face25ada-datadruids /movie-metadatatabular10K<n<100K0 likes37 downloads2y agoHugging Face26moviebrain01 /anime-dataset-2025 Anime Dataset 2025 Description This dataset contains anime metadata used for machine learning and recommendation systems. Splits train test Columns Includes anime title, genres, score, members and other metadata. Use cases Anime recommendation systems NLP tasks Machine learning projects License CC-BY-4.0 image10K<n<100K1 likes36 downloads7mo agoHugging Face27DropTheHQ /movie-ratings Movie Ratings Database 154,965 movies with ratings, vote counts, release dates, languages, genres, and runtime. Source DropThe.org — Data platform tracking 209K+ movies. Analysis 209K Movies Ratings Analysis 191K Movies Feelgood Score Links DropThe.org Movie Statistics Methodology tabulartabular-regression100K<n<1M0 likes35 downloads7mo agoHugging Face28SandipPalit /Movie_Datasettabulartext-classification10K<n<100K8 likes33 downloads4y agoHugging Face29saikiranmaddukuri /moviedbtabular10K<n<100K0 likes31 downloads25d agoHugging Face30jason1966 /ahsanaseer_top-rated-tmdb-movies-10k TMDB Movies Dataset Dataset of 10k top rated TMDB movies for text preprocessing (NLP) Dataset Info Source: Kaggle Original Size: 1.43 MB Kaggle Downloads: 8,035 Files: 1 Files top10K-TMDB-movies.csv Mirrored from Kaggle tabular10K<n<100K0 likes30 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.