datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
recommendationhnm-fashion-recommendations-data
Dataset Rekomendasi Fashion H&M
Dataset ini berisi data transaksi, atribut pelanggan, dan metadata produk yang telah dianonimkan dari H&M Group. Kumpulan data komprehensif ini memungkinkan pemodelan perilaku pembelian pelanggan secara mendalam.
Wawasan yang dihasilkan dapat dimanfaatkan untuk berbagai tujuan bisnis yang strategis, mulai dari meningkatkan personalisasi pengalaman berbelanja, mengoptimalkan manajemen inventaris untuk efisiensi produksi, hingga mendukung inisiatif… See the full description on the dataset page: https://huggingface.co/datasets/einrafh/hnm-fashion-recommendations-data.zomato-restaurant-recommendationpaper-recommendations-v2aksahaha_crop-recommendation
crop recommendation
Crop Growth Recommendations: Optimal Conditions for Higher Yields
Dataset Info
Source: Kaggle
Original Size: 0.06 MB
Kaggle Downloads: 4,065
Files: 1
Files
Crop_recommendation.csv
Mirrored from Kaggle
fashion-recommendation-datasetCrop-recommendationCrop-Recommendation-Parameters
🌱 Crop Recommendation Dataset
A machine learning dataset for crop recommendation based on soil properties and environmental conditions. The dataset contains measurements of essential soil nutrients and climatic parameters, along with the crop label that is suitable for those conditions.
This dataset can be used for machine learning classification, agricultural analytics, decision-support systems, and smart farming applications.
📌 Dataset Overview
Property… See the full description on the dataset page: https://huggingface.co/datasets/Samarth-27/Crop-Recommendation-Parameters.synthesized-cloud-optimization-recommendations
Synthesized Cloud-Optimization Recommendations
18 scenarios that pair cloud telemetry with a hand-crafted optimization
recommendation. Use them to train models or to evaluate AI agents.
Summary
Each scenario has multi-tier telemetry, a Terraform file describing the
deployed infrastructure, and a gold-standard recommendation.
The dataset is built around a simple input-output mapping. The input is
telemetry plus the infrastructure. The output is an optimization… See the full description on the dataset page: https://huggingface.co/datasets/ameau01/synthesized-cloud-optimization-recommendations.HM-Personalized-Fashion-Recommendationsfashion-recommendation-images
High-Resolution Fashion Product Images
This dataset is a highly optimized, high-resolution subset of the popular Fashion Product Images Dataset originally hosted on Kaggle.
It contains thousands of unique e-commerce fashion products, combining high-resolution product images with multiple descriptive label attributes.
All low-resolution thumbnails and anomalies have been aggressively filtered out. Every image in this dataset has a minimum resolution of 640px on its shortest… See the full description on the dataset page: https://huggingface.co/datasets/GangHitman/fashion-recommendation-images.HuggingBench-Recommendation
HuggingBench-Recommendation
This dataset contains Resource Recommendation test collection in HuggingBench for paper "Benchmarking Recommendation, Classification, and Tracing Based on Hugging Face Knowledge Graph".
Dataset Details
general_rec contains training/validation/test set files for the General Collaborative Filtering methods in the format required by SSLRec.
social_rec contains training/validation/test set files and user social relation file for the Social… See the full description on the dataset page: https://huggingface.co/datasets/cqsss/HuggingBench-Recommendation.movie_recommendationMovie recommendation task based on the Movielens datasetBangla-Book-Recommendation-Dataset
Summary
This repository contains the dataset for the paper Towards Personalized Bangla Book Recommendation: A Large-Scale Multi-Entity Book Graph Dataset. In this work, we introduce a large-scale multi-entity heterogeneous graph dataset for Bangla book recommendation, integrating users, books, authors, publishers, and categories. Our framework enables the development of sophisticated graph-based recommendation systems. This contribution aims to enhance the discovery of Bangla… See the full description on the dataset page: https://huggingface.co/datasets/DevnilMaster1/Bangla-Book-Recommendation-Dataset.music-recommendations
Group aggregators over a frozen scorer — artifacts (YAMBDA-50m)
Чекпоинты и промежуточные артефакты для групповых музыкальных рекомендаций:
per-user скорер, кэш его топ-200, аудиоэмбеддинги каталога и обученные
групповые агрегаторы. Всё производное от
yandex/yambda, flavor 50m.
Состав
Путь
Что это
gsasrec/best.pt, config.json, metrics.csv
SASRec-скорер, 276 305 items, test NDCG@10 = 0.0726
gsasrec/item_id_to_idx.pkl
item_id YAMBDA → компактный индекс… See the full description on the dataset page: https://huggingface.co/datasets/Vladislavbro-500/music-recommendations.Fashion-Recommendation-Images
Fashion Recommendation Images
Cleaned fashion image dataset used for the Match AI Fashion Recommendation System.
Dataset
31,377 cleaned fashion images
Duplicate, corrupted and placeholder images removed
Images used for CLIP feature extraction and visual similarity matching
Original Dataset
Vibrent Clothes Rental DatasetSource: Kagglehttps://www.kaggle.com/datasets/kaborg15/vibrent-clothes-rental-dataset
Usage
The images are used to… See the full description on the dataset page: https://huggingface.co/datasets/aishwaryas61622/Fashion-Recommendation-Images.product-recommendation-2025
TREC 2025 Product Recommendation Data
This is the data for the recommendation task for the TREC 2025 Product Search and Recommendation Task.
The initial directory contains the initial corpus and training data release.
This may be updated as we get further along in the timeline.
[!NOTE]
This data is derived from the Amazon ESCI and M2 data sets, each under the
Apache license (version 2.0).
Yelp-Multimodal-Recommendation
Yelp-MultimodalRec
A multimodal dataset for POI (Point of Interest) recommendation, based on the Yelp Open Dataset.It includes business metadata, user reviews, business photos, and LLM-generated summaries of reviews and images.This dataset supports downstream tasks like session-based recommendation, multimodal embedding learning, and more.
📁 Dataset Structure
File Name
Description
business.csv
Business metadata including name, address, categories, etc.… See the full description on the dataset page: https://huggingface.co/datasets/bridgekk/Yelp-Multimodal-Recommendation.chat_restaurant_recommendation
Restaurant chat dataset
This dataset contains approximately 600 chat interactions mimicking various user tones with different restaurant categories
The data were generated by Gemini-pro
The purpose of this dataset is to serve as a calibration dataset for a restaurant recommendation LLM chatbot
visual-product-recommendations-catalogueSupport-Bot-Recommendationrecommendation-engine-dataAmazon-Reviews-2023-Recommendationsynthetic-product-recommendation-examples
Synthetic Product Recommendation Examples
An entirely synthetic bilingual dataset of ecommerce discovery queries paired with candidate products and graded relevance judgments. It is designed for educational retrieval, reranking, and recommendation experiments and contains no private catalog, merchant, customer, behavioral, or transaction data.
Dataset Description
The dataset contains twenty English and French queries with ten candidates per query. It complements… See the full description on the dataset page: https://huggingface.co/datasets/neurocheckout-ai/synthetic-product-recommendation-examples.tcfd_recommendations
Dataset Card for tcfd_recommendations
Dataset Summary
We introduce an expert-annotated dataset for classifying the TCFD recommendation categories (fsb-tcfd.org) of paragraphs in corporate disclosures.
Supported Tasks and Leaderboards
The dataset supports a multiclass classification task of paragraphs into the four TCFD recommendation categories (governance, strategy, risk management, metrics and targets) and the non-climate-related class.
Languages… See the full description on the dataset page: https://huggingface.co/datasets/climatebert/tcfd_recommendations.authority-bias-paper-recommendation
Authority Bias in Conversational Search Engines for Academic Paper Recommendation
Dataset accompanying the paper "Authority Bias in Conversational Search Engines for Academic Paper Recommendation" (EMNLP 2026, Main Conference). Code: https://github.com/jinaduuthman/Authority-Bias-In-Conversational-Search-Engine
This is a content-controlled counterfactual audit of authority bias in LLM paper recommendation. Each paper's content (title + abstract) is held fixed while its authority… See the full description on the dataset page: https://huggingface.co/datasets/uthmanjinadu/authority-bias-paper-recommendation.Yelp-Multimodal-Recommendation
Yelp-MultimodalRec
A multimodal dataset for POI (Point of Interest) recommendation, based on the Yelp Open Dataset.It includes business metadata, user reviews, business photos, and LLM-generated summaries of reviews and images.This dataset supports downstream tasks like session-based recommendation, multimodal embedding learning, and more.
📁 Dataset Structure
File Name
Description
business.csv
Business metadata including name, address, categories… See the full description on the dataset page: https://huggingface.co/datasets/conananywhere/Yelp-Multimodal-Recommendation.Yelp-Multimodal-Recommendation
Yelp-MultimodalRec
A multimodal dataset for POI (Point of Interest) recommendation, based on the Yelp Open Dataset.It includes business metadata, user reviews, business photos, and LLM-generated summaries of reviews and images.This dataset supports downstream tasks like session-based recommendation, multimodal embedding learning, and more.
📁 Dataset Structure
File Name
Description
business.csv
Business metadata including name, address, categories, etc.… See the full description on the dataset page: https://huggingface.co/datasets/wzehui/Yelp-Multimodal-Recommendation.Road_construction_recommendationLLM_for_Dietary_Recommendation_System
Chat_GPT_for_Nutritional_Recommendation_System
This repository contains the code, 50 different patient profiles, and respective Chat-GPT responses with nutritional recommendations and sample diet plans. Patient profiles contain easy, medium, and complex cases and various diseases. The project aims to evaluate the application of large language models for nutritional recommendation systems.
Responses were evaluated based on personalization, consistency with evidence-based… See the full description on the dataset page: https://huggingface.co/datasets/issai/LLM_for_Dietary_Recommendation_System.
