datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
waste-classification
Waste Classification (repackaged)
Summary:
A repackaged version of the Kaggle “Waste Classification” dataset with a consistent multi-choice training schema and multiple splits.
Splits:
cleaned: Only real-world photos that match the declared subclass (non-photos, PPT slides, icons, cartoons, or mismatches removed).
Schema (columns):
image: Image file (datasets.Image).
class: One of the four top-level categories.
subclass: Fine-grained category (from folder… See the full description on the dataset page: https://huggingface.co/datasets/huaweilin/waste-classification.waste-classification-audio-helsinki
Dataset Card for "waste-classification-audio"
english to italian translation was made with helsinki-NLP translation model.
More Information needed
waste-classification-v3waste-classificationwaste-classification-v2
Dataset Card for Dataset Name
Dataset Summary
Dataset used to train a language model to do classification on 50 different waste classes.
Languages
English
Dataset Structure
Data Instances
Phrase
Class
Index
"I have this apple phone charger to throw, where should I put it ?"
PHONE CHARGER
26
"Should I recycle a disposable cup ?"
Plastic Cup
32
"I have a milk brick"
Tetrapack
45
Data Fields
Phrase
Class… See the full description on the dataset page: https://huggingface.co/datasets/thomasavare/waste-classification-v2.waste-classification-audio-deepl-largeconcatentation of "thomasavare/waste-classification-audio-deepl" and "thomasavare/waste-classification-audio-deepl2", so its easier to use.
there are 3 duplicates, idk why but it's too long to remake the datasets (maybe later) and not necessary.
waste-classification-audio-helsinki2
Dataset Card for "waste-classification-audio-helsinki2"
More Information needed
waste-classification-audio-deepl2audio created from dataset "italian-dataset-deepl2"
waste-classification-audio-deepl
Dataset Card for "waste-classification-audio-deepl"
More Information needed
waste-classification-audio-deepl-v3waste-classification-2waste-classification-unseen-gptaugmented
dataset_info:
features:
- name: original_id
dtype: int64
- name: Phrase
dtype: string
- name: Class_index
dtype: float64
splits:
- name: train
num_bytes:
num_examples: 1000
download_size:
dataset_size:
configs:
- config_name: default
data_files:
- split: train
path: data/train-*
chaptgpt modified dataset from unused phrases from thomasavare-waste-classification-v2waste-classification-audio-unseen-gpt
