autonlp
Datasets
All datasets matching “autonlp”autonlp-data-song-lyrics
Song Lyrics Genre Classification
A dataset of 53,882 song lyrics labeled across 6 musical genres, used to train the juliensimon/autonlp-song-lyrics-18753417 genre classifier.
Video walkthrough: Classifying song lyrics with Hugging Face AutoNLP
Dataset Details
Detail
Value
Task
Multi-class text classification (6 genres)
Language
English
Total samples
53,882
Processing
Hugging Face AutoNLP
Genre Labels
ID
Genre
0
Dance
1
Heavy… See the full description on the dataset page: https://huggingface.co/datasets/juliensimon/autonlp-data-song-lyrics.autonlp-data-antisemitism-2
AutoNLP Dataset for project: antisemitism-2
Table of content
Dataset Description
Languages
Dataset Structure
Data Instances
Data Fields
Data Splits
Dataset Descritpion
This dataset has been automatically processed by AutoNLP for project antisemitism-2.
Languages
The BCP-47 code for the dataset's language is en.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"target":… See the full description on the dataset page: https://huggingface.co/datasets/astarostap/autonlp-data-antisemitism-2.autonlp-data-user-review-classification
AutoNLP Dataset for project: user-review-classification
Table of content
Dataset Description
Languages
Dataset Structure
Data Instances
Data Fields
Data Splits
Dataset Descritpion
This dataset has been automatically processed by AutoNLP for project user-review-classification.
Languages
The BCP-47 code for the dataset's language is en.
Dataset Structure
Data Instances
A sample from this dataset looks as… See the full description on the dataset page: https://huggingface.co/datasets/alperiox/autonlp-data-user-review-classification.autonlp-data-second
AutoNLP Dataset for project: second
Table of content
Dataset Description
Languages
Dataset Structure
Data Instances
Data Fields
Data Splits
Dataset Descritpion
This dataset has been automatically processed by AutoNLP for project second.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"text": "one hundred and… See the full description on the dataset page: https://huggingface.co/datasets/VoidZeroe/autonlp-data-second.autonlp-data-peptidesDeep learning the collisional cross sections of the peptide universe from a million experimental values
Data generated from MaxQuant output
wget https://ftp.pride.ebi.ac.uk/pride/data/archive/2020/12/PXD017703/HeLa_200ng_Library_MaxQuant.zip
unzip HeLa_200ng_Library_MaxQuant.zip
awk -F '\t' '{print $1,",",$40}' evidence.txt > pepCCS.csv
wc pepCCS.csv
352111 1056333 12736697 pepCCS.csv
Code
autonlp-data-Scientific_Title_Generator
AutoNLP Dataset for project: Scientific_Title_Generator
Table of content
Dataset Description
Languages
Dataset Structure
Data Instances
Data Fields
Data Splits
Dataset Descritpion
This dataset has been automatically processed by AutoNLP for project Scientific_Title_Generator.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as… See the full description on the dataset page: https://huggingface.co/datasets/AryanLala/autonlp-data-Scientific_Title_Generator.
