datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CelebA-attrs
CelebA-128x128
CelebA with attrs at 128x128 resolution.
Dataset Information
The attributes are binary attributes. The dataset is already split into train/test/validation sets.
Citation
@inproceedings{liu2015faceattributes,
title = {Deep Learning Face Attributes in the Wild},
author = {Liu, Ziwei and Luo, Ping and Wang, Xiaogang and Tang, Xiaoou},
booktitle = {Proceedings of International Conference on Computer Vision (ICCV)},
month = {December},
year… See the full description on the dataset page: https://huggingface.co/datasets/tpremoli/CelebA-attrs.bone_marrow_cell_dataset
About This Dataset
Bone marrow biopsy is procedure applied to collect and examine bone marrow — the spongy tissue inside some of your larger bones.
This biopsy can show whether your bone marrow is healthy and making normal amounts of blood cells. Doctors use these procedures to diagnose and monitor blood and marrow diseases, cancers, as well as fevers of unknown origin.
The dataset contains a collection of over 170,000 de-identified, expert-annotated cells from the bone marrow… See the full description on the dataset page: https://huggingface.co/datasets/ekim15/bone_marrow_cell_dataset.DigiCam-CelebA-26KData is measured at 30 cm, as shown below.
After downloading and installing LenslessPiCam, the simulated PSF can be obtained and compared with the measured one with the following command:
python scripts/sim/digicam_psf.py \
huggingface_repo=bezzam/DigiCam-CelebA-26K \
sim.waveprop=False \
sim.deadspace=True \
digicam.gamma=2.2 \
digicam.ap_center="[58,76]" \
digicam.ap_shape="[19,25]" \
digicam.rotate=0 \
digicam.horizontal_shift=-60 \
digicam.vertical_shift=-80
For a… See the full description on the dataset page: https://huggingface.co/datasets/bezzam/DigiCam-CelebA-26K.DigiCam-CelebA-10K
Dataset for the paper: https://opg.optica.org/abstract.cfm?uri=pcAOP-2023-JTu4A.45
Data is measured with a computer monitor at 30 cm as shown below (except for the in-the-wild mug measurement which is measured at 12 cm).
After cloning and installing LenslessPiCam, ADMM reconstruction can be applied to the dataset with this script (handles dataset downloading from Hugging Face).python scripts/recon/dataset.py -cn recon_celeba_digicam
The simulated PSF can be obtained and compared with the… See the full description on the dataset page: https://huggingface.co/datasets/bezzam/DigiCam-CelebA-10K.metacloak_celeba_vggface2
Dataset Card for MetaCloak
Dataset Summary
This repository provides datasets from the MetaCloak.
For each dataset, *-gen is the subset used for protecting, and *-eval is used as a clean reference to calculate some quality metrics.
from datasets import load_dataset
dataset = load_dataset("yixin/metacloak_celeba_vggface2")
Contact
Contact Us: yixinliucs@gmail.com
celeb-fbi
Celeb-FBI: Celebrity Full Body Images Dataset
A cleaned and restructured version of the Celeb-FBI dataset containing 7,208 full-body celebrity images with annotations for height, weight, age, and gender.
Dataset Description
This dataset consists of worldwide celebrity images captured in standing, front-facing positions. It is designed for research on human attribute estimation from full-body images, including height, weight, age, and gender prediction tasks.… See the full description on the dataset page: https://huggingface.co/datasets/alecccdd/celeb-fbi.celebahq_512_id_clusters
celebahq_512 with SRK identity labels
Summary
This dataset is a derived version of jxie/celeba-hq. It keeps the original image set and adds automatically generated identity-group labels derived from face-embedding clustering.
As explained in our experimental setup, we use CelebA-HQ from Karras et al. (2018), specifically the Hugging Face snapshot at revision 7ecc6a45edfb5483ccf2f7df1035d298ffe7c76b. The referenced CelebA-HQ version provides gender labels but no identity… See the full description on the dataset page: https://huggingface.co/datasets/edgarcancinoe/celebahq_512_id_clusters.face-celeb-vietnamese
Dataset Card for "face-celeb-vietnamese"
Dataset Summary
This dataset contains information on over 8,000 samples of well-known Vietnamese individuals, categorized into three professions: singers, actors, and beauty queens. The dataset includes data on more than 100 celebrities in each of the three job categories.
Languages
Vietnamese: The label is used to indicate the name of celebrities in Vietnamese.
Dataset Structure
The image and Vietnamese… See the full description on the dataset page: https://huggingface.co/datasets/fptudsc/face-celeb-vietnamese.celeb-fbi
Celeb-FBI: Celebrity Full Body Images Dataset
A cleaned and restructured version of the Celeb-FBI dataset containing 7,208 full-body celebrity images with annotations for height, weight, age, and gender.
Dataset Description
This dataset consists of worldwide celebrity images captured in standing, front-facing positions. It is designed for research on human attribute estimation from full-body images, including height, weight, age, and gender prediction tasks.… See the full description on the dataset page: https://huggingface.co/datasets/Dongrae/celeb-fbi.celeb-fbi-pose-estimation
Celeb-FBI Pose Estimation Dataset
A derivative of the Celeb-FBI dataset enriched with 3D human pose and mesh recovery annotations generated using Meta's SAM 3D Body model.
Dataset Description
This dataset extends the original Celeb-FBI celebrity full-body image dataset with comprehensive 3D pose estimation outputs. Each sample includes predicted 3D keypoints, 2D keypoints, mesh vertices, body/hand pose parameters, and shape parameters extracted using the… See the full description on the dataset page: https://huggingface.co/datasets/alecccdd/celeb-fbi-pose-estimation.vietnam_celeb_faceCeline.Product.prices.Germany
Celine web scraped data
About the website
Celine operates within the fashion industry in the EMEA region, particularly in Germany. This industry is currently undergoing digital transformation with a focus on Ecommerce, creating a competitive space for established and emerging fashion brands. In Germany, the local customer using various online stores has shown significant growth, displaying an increased interest in online fashion shopping. This pattern presents the fashion… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Celine.Product.prices.Germany.Celine.Product.prices.United.Kingdom
Celine web scraped data
About the website
The Ecommerce industry in the EMEA region, particularly in the United Kingdom, has shown a significant growth rate in recent years, becoming a pivotal element in the regions economic landscape. The rise of digital technologies and change in consumer behavior has further accelerated this upward trend. In this digital marketplace, Celine, a well-known high-end fashion brand, has maintained its mark. The dataset at-hand encompasses… See the full description on the dataset page: https://huggingface.co/datasets/DBQ/Celine.Product.prices.United.Kingdom.celeb-fbi
Celeb-FBI: Celebrity Full Body Images Dataset
A cleaned and restructured version of the Celeb-FBI dataset containing 7,208 full-body celebrity images with annotations for height, weight, age, and gender.
Dataset Description
This dataset consists of worldwide celebrity images captured in standing, front-facing positions. It is designed for research on human attribute estimation from full-body images, including height, weight, age, and gender prediction tasks.… See the full description on the dataset page: https://huggingface.co/datasets/farsee04/celeb-fbi.oxford-pets
Oxford-IIIT Pet Dataset
Images from The Oxford-IIIT Pet Dataset. Only images and labels have been pushed, segmentation annotations were ignored.
Homepage: https://www.robots.ox.ac.uk/~vgg/data/pets/
License:
Same as the original dataset.
face-celeb-vietnamese
Dataset Card for "face-celeb-vietnamese"
Dataset Summary
This dataset contains information on over 8,000 samples of well-known Vietnamese individuals, categorized into three professions: singers, actors, and beauty queens. The dataset includes data on more than 100 celebrities in each of the three job categories.
Languages
Vietnamese: The label is used to indicate the name of celebrities in Vietnamese.
Dataset Structure
The image and Vietnamese… See the full description on the dataset page: https://huggingface.co/datasets/thuanML/face-celeb-vietnamese.
