datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
imagenet1k-256-wdsThis is imagenet1k in webdataset format. Images are stored as jpg files. Every image has been resized to a maximum side length of 256. That means that if an image in the original dataset was 1000 by 500, the new size will be 256 by 128. Images with a maximum side length of under 256 were not resized.
The total size of all dataset files is 57.8 GB, there are 1,281,167 rows in the training split and 50,000 rows in the validation split.
autotrain-data-attempt
AutoTrain Dataset for project: attempt
Dataset Description
This dataset has been automatically processed by AutoTrain for project attempt.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"image": "<800x1000 RGB PIL image>",
"target": 13
},
{
"image": "<254x512 RGB PIL image>",
"target": 0
}]
Dataset Fields
The… See the full description on the dataset page: https://huggingface.co/datasets/AdamOswald1/autotrain-data-attempt.autotrain-data-alt
AutoTrain Dataset for project: alt
Dataset Description
This dataset has been automatically processed by AutoTrain for project alt.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"image": "<600x600 RGB PIL image>",
"target": 1
},
{
"image": "<1024x590 RGB PIL image>",
"target": 1
}]
Dataset Fields
The dataset… See the full description on the dataset page: https://huggingface.co/datasets/AdamOswald1/autotrain-data-alt.yelp-04-2024
Yelp Complete Open Dataset 04.2024
Dataset Description
This dataset contains the complete Yelp Open Dataset, a rich collection of user reviews, business information, and user data. It is a valuable resource for tasks such as sentiment analysis, recommendation systems, and other natural language processing (NLP) projects.
Source
The dataset is provided by Yelp and is publicly available under the Yelp Dataset Terms of Use.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/adamamer20/yelp-04-2024.grape-leaf-disease-augmented-dataset
🍇🍃 Grape Leaf Disease Augmented Dataset`
Information
This dataset consists of 18,054 grape leaf images across 4 class (ESCA, Leaf Blight, Black Rot, Healthy) for use in image classification tasks.
Dataset has been created as a modification to the original dataset from the PlantVillage dataset: https://www.kaggle.com/datasets/abdallahalidev/plantvillage-dataset
The images have been used from this dataset:… See the full description on the dataset page: https://huggingface.co/datasets/adamkatchee/grape-leaf-disease-augmented-dataset.yolo-emotionsA merged emotions dataset was created using a highly curated subset of ExpW, FER2013 (enhanced with FER2013+), AffectNet (6 emotions), and RAF-DB in YOLO format, totaling approximately 155K samples. A YOLOv11-x model, fine-tuned on the WiderFace dataset for the bounding boxes, was used. The distribution is as follows:
TRAIN Set Class Distribution:
Class 0 (Angry): 8511 (6.84%)
Class 1 (Disgust): 6307 (5.07%)
Class 2 (Fear): 4249 (3.41%)
Class 3 (Happy): 37714 (30.30%)
Class 4… See the full description on the dataset page: https://huggingface.co/datasets/AdamCodd/yolo-emotions.autotrain-data-let
AutoTrain Dataset for project: let
Dataset Description
This dataset has been automatically processed by AutoTrain for project let.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"image": "<600x600 RGB PIL image>",
"target": 1
},
{
"image": "<1024x590 RGB PIL image>",
"target": 1
}]
Dataset Fields
The dataset… See the full description on the dataset page: https://huggingface.co/datasets/AdamOswald1/autotrain-data-let.
