CoolFace
Datasetpublic

akaruineko/fantastic-offensive

fantastic-offensive dataset Binary offensive-language classification dataset combining several public sources and augmented with an obfuscation engine (leet-speak, separators, censoring, repeated chars, case shuffle, unicode homoglyphs, fullwidth) so classifiers learn to detect censored / mutated curse words. Schema column type meaning text string input sentence label int8 1 = offensive, 0 = clean source string originating dataset origin_label… See the full description on the dataset page: https://huggingface.co/datasets/akaruineko/fantastic-offensive.

sourceHugging Facecc0-1.0updated 1mo agoView on Hugging Face
0likes77downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
akaruineko/fantastic-offensive · CoolFace