inctdd/autotrain-data-told_br_binary_sm_bertimbau
AutoTrain Dataset for project: told_br_binary_sm_bertimbau Dataset Description This dataset has been automatically processed by AutoTrain for project told_br_binary_sm_bertimbau. Languages The BCP-47 code for the dataset's language is unk. Dataset Structure Data Instances A sample from this dataset looks as follows: [ { "text": "@user agora n\u00e3o me d\u00e1 mais, mas antes, porra", "target": 1 }, {… See the full description on the dataset page: https://huggingface.co/datasets/inctdd/autotrain-data-told_br_binary_sm_bertimbau.
AutoTrain Dataset for project: toldbrbinarysmbertimbau
Dataset Description
This dataset has been automatically processed by AutoTrain for project toldbrbinarysmbertimbau.
Languages
The BCP-47 code for the dataset's language is unk.
Dataset Structure
Data Instances
A sample from this dataset looks as follows:
[
{
"text": "@user agora n\u00e3o me d\u00e1 mais, mas antes, porra",
"target": 1
},
{
"text": "pires \u00e9 fodido fds mais um",
"target": 1
}
]Dataset Fields
The dataset has the following fields (also called "features"):
{
"text": "Value(dtype='string', id=None)",
"target": "ClassLabel(names=['0', '1'], id=None)"
}Dataset Splits
This dataset is split into a train and validation split. The split sizes are as follow:
