datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
viet-cultural-vqaVietnamese Cultural VQA Dataset is a comprehensive multimodal dataset focusing on Vietnamese cultural heritage.
It contains 28,505 images across 12 cultural categories with 119,012 question-answer pairs in Vietnamese and English.
The dataset covers diverse aspects of Vietnamese culture including architecture, cuisine, traditional clothing,
landscapes, festivals, folk culture, traditional games, sports, handicrafts, musical instruments, daily life,
and transportation.viet-cultural-vqaVietnamese Cultural VQA Dataset is a comprehensive multimodal dataset focusing on Vietnamese cultural heritage.
It contains 28,505 images across 12 cultural categories with 119,012 question-answer pairs in Vietnamese and English.
The dataset covers diverse aspects of Vietnamese culture including architecture, cuisine, traditional clothing,
landscapes, festivals, folk culture, traditional games, sports, handicrafts, musical instruments, daily life,
and transportation.viet-cultural-vqa
🇻🇳 Vietnamese Cultural VQA Dataset
📖 Dataset Description
The Vietnamese Cultural VQA Dataset is a comprehensive multimodal dataset designed for Visual Question Answering (VQA) tasks focused on Vietnamese cultural heritage. This dataset aims to bridge the gap in understanding and preserving Vietnamese culture through AI-powered visual understanding and question answering.
🎯 Dataset Summary
📊 Total Images: 28,505 high-quality cultural images
💬 Total… See the full description on the dataset page: https://huggingface.co/datasets/IAmFuch/viet-cultural-vqa.ego4d-random-views-20k
Ego4D Random Views Dataset
This dataset contains 20,000 random view frames sampled from the Ego4D dataset using a high-performance multi-process generation system.
Dataset Overview
Total Images: 20,000 high-quality frames
Image Format: PNG (1024×1024 resolution)
Source: Ego4D v2 dataset (52,665+ video files)
Sampling Method: Multi-process random sampling with maximum diversity
Generation Time: 797.57 seconds (~13 minutes)
Generation Speed: 25.08 frames/second… See the full description on the dataset page: https://huggingface.co/datasets/weikaih/ego4d-random-views-20k.cardiac_cine_acdc
ACDC (Cardiac Cine-MRI)
ACDC (Automatic Cardiac Diagnosis Challenge, MICCAI 2017) is a cine‑MRI dataset for cardiac segmentation.This repository contains processed NIfTI files in Data/processed_output/acdc format.
Dataset Summary
Modality: Cardiac cine‑MRI (NIfTI)
Task: Segmentation of LV, RV, and myocardium
Frames: ED/ES + full SAX time series (sax_t)
Labels: LV/RV cavities + myocardium
Splits: train, test (as provided in processed output)
Data Structure (per… See the full description on the dataset page: https://huggingface.co/datasets/viennh2012/cardiac_cine_acdc.viewpoint-aware-pig-posture-recognition
Viewpoint-Aware Pig Posture Recognition Dataset
This dataset supports multi-camera, viewpoint-aware pig posture recognition in livestock barn environments. It contains real-world pig images, bounding box annotations, posture class labels, and per-instance camera viewpoint angles (azimuth and elevation) derived from PnP-based camera calibration.
Code: Anil-Bhujel/viewpoint-aware-pig-posture-recognition on GitHub
Dataset Summary
Images were captured from 2… See the full description on the dataset page: https://huggingface.co/datasets/anilbhujel/viewpoint-aware-pig-posture-recognition.face-celeb-vietnamese
Dataset Card for "face-celeb-vietnamese"
Dataset Summary
This dataset contains information on over 8,000 samples of well-known Vietnamese individuals, categorized into three professions: singers, actors, and beauty queens. The dataset includes data on more than 100 celebrities in each of the three job categories.
Languages
Vietnamese: The label is used to indicate the name of celebrities in Vietnamese.
Dataset Structure
The image and Vietnamese… See the full description on the dataset page: https://huggingface.co/datasets/fptudsc/face-celeb-vietnamese.generated-vietnamese-passeports-datasetData generation in machine learning involves creating or manipulating data to train
and evaluate machine learning models. The purpose of data generation is to provide
diverse and representative examples that cover a wide range of scenarios, ensuring the
model's robustness and generalization.
The dataset contains GENERATED Vietnamese passports, which are replicas of official
passports but with randomly generated details, such as name, date of birth etc.
The primary intention of generating these fake passports is to demonstrate the
structure and content of a typical passport document and to train the neural network to
identify this type of document.
Generated passports can assist in conducting research without accessing or compromising
real user data that is often sensitive and subject to privacy regulations. Synthetic
data generation allows researchers to *develop and refine models using simulated
passport data without risking privacy leaks*.vietnam_celeb_faceviewer-test-large-images-fail
Test Dataset: viewer-test-large-images-fail
Purpose: Demonstrate dataset viewer behavior with large images and row group sizing.
Dataset Details
Number of images: 50
Total size: ~632MB
Average per image: ~12.6MB
Expected Behavior
50 large images (~14MB each, ~700MB total). With 100-row row groups, all images land in one ~700MB row group, exceeding the 300MB scan limit.
Prediction: FAIL - Scan size limit exceeded
Technical Context
The dataset viewer… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/viewer-test-large-images-fail.Vietnamese_StreetFood_Dataset
Vietnamese Street Food Dataset
This dataset contains approximately 3,500 real-world images of popular Vietnamese street food dishes. The images were collected to support tasks such as image classification, food recognition, and cultural heritage preservation related to Vietnamese cuisine.
All photos focus on authentic street-style presentations — from markets, vendors, and everyday eating spots across Vietnam. The dataset is particularly useful for training models to identify… See the full description on the dataset page: https://huggingface.co/datasets/HoagMin/Vietnamese_StreetFood_Dataset.durian_disease_classification_vietnam
Durian Disease Classification Vietnam
A dataset for image classification of Durian Disease. The dataset contains 2,595 images across 6 classes: Leaf_Algal, Leaf_Blight, Leaf_Colletotrichum, Leaf_Healthy, Leaf_Phomopsis, Leaf_Rhizoctonia.
Images per class:
Leaf_Algal: 462
Leaf_Blight: 440
Leaf_Colletotrichum: 400
Leaf_Healthy: 484
Leaf_Phomopsis: 411
Leaf_Rhizoctonia: 398
This dataset is indexed on https://project-agml.github.io/ as part of the AgML python library.… See the full description on the dataset page: https://huggingface.co/datasets/Project-AgML/durian_disease_classification_vietnam.face-celeb-vietnamese
Dataset Card for "face-celeb-vietnamese"
Dataset Summary
This dataset contains information on over 8,000 samples of well-known Vietnamese individuals, categorized into three professions: singers, actors, and beauty queens. The dataset includes data on more than 100 celebrities in each of the three job categories.
Languages
Vietnamese: The label is used to indicate the name of celebrities in Vietnamese.
Dataset Structure
The image and Vietnamese… See the full description on the dataset page: https://huggingface.co/datasets/thuanML/face-celeb-vietnamese.viewer-test-50-small-images-ok
Test Dataset: viewer-test-50-small-images-ok
Purpose: Demonstrate dataset viewer behavior with large images and row group sizing.
Dataset Details
Number of images: 50
Total size: ~167MB
Average per image: ~3.3MB
Expected Behavior
50 smaller images (~3.5MB each, ~175MB total). Well under the 300MB limit.
Prediction: SHOULD WORK - Under 300MB limit
Technical Context
The dataset viewer converts imagefolder datasets to parquet with:
Row group size:… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/viewer-test-50-small-images-ok.viewer-test-25-large-images-fail
Test Dataset: viewer-test-25-large-images-fail
Purpose: Demonstrate dataset viewer behavior with large images and row group sizing.
Dataset Details
Number of images: 25
Total size: ~318MB
Average per image: ~12.7MB
Expected Behavior
25 large images (~14MB each, ~350MB total). Even with fewer images, the row group still exceeds 300MB.
Prediction: FAIL - Scan size limit exceeded
Technical Context
The dataset viewer converts imagefolder datasets to… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/viewer-test-25-large-images-fail.viewer-test-15-large-images-ok
Test Dataset: viewer-test-15-large-images-ok
Purpose: Demonstrate dataset viewer behavior with large images and row group sizing.
Dataset Details
Number of images: 15
Total size: ~188MB
Average per image: ~12.6MB
Expected Behavior
15 large images (~14MB each, ~210MB total). Under the 300MB limit, so the viewer should work.
Prediction: SHOULD WORK - Under 300MB limit
Technical Context
The dataset viewer converts imagefolder datasets to parquet with:… See the full description on the dataset page: https://huggingface.co/datasets/davanstrien/viewer-test-15-large-images-ok.PromptedArtistIdentificationDataset-ViewSamples
Sample Dataset Viewer for Prompted Artist Identification Dataset
Website | Paper | GitHub
Identifying Prompted Artist Names from Generated Images
Grace Su, Sheng-Yu Wang, Aaron Hertzmann, Eli Shechtman, Jun-Yan Zhu, Richard Zhang
arXiv, 2025
Description
This page serves as a viewer for sample images from the Prompted Artist Identification Dataset.
Please visit the main dataset page for a description of the full dataset.
The entire benchmark dataset consists of 1.95… See the full description on the dataset page: https://huggingface.co/datasets/cmu-gil/PromptedArtistIdentificationDataset-ViewSamples.vietnamese-food-images
vietnamese-food-images
Food images collected from Google Maps restaurant reviews, with rich metadata
(place, location, review, dish classification, image-quality scores).
Built with the pipeline in
crawl_image — crawl review photos → AI food/dish
classification + quality filtering → upload.
Statistics
Images: 34104
Unique places: 1989
Dish classes: 33
Dish distribution
dish
count
other_dish
5738
bun_dau_mam_tom
4148
lau
2710… See the full description on the dataset page: https://huggingface.co/datasets/ThanThoai9x/vietnamese-food-images.
