CoolFace
2 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Ganasekhar /pii-masking-400k Purpose and Features 🌍 World's largest open dataset for privacy masking 🌎 The dataset is useful to train and evaluate models to remove personally identifiable and sensitive information from text, especially in the context of AI assistants and LLMs. AI4Privacy Dataset Analytics 📊 Dataset Overview Total entries: 406,896 Total tokens: 20,564,179 Total PII tokens: 2,357,029 Number of PII classes in public dataset: 17 Number of PII classes in extended dataset:… See the full description on the dataset page: https://huggingface.co/datasets/Ganasekhar/pii-masking-400k.texttext-classification100K<n<1M0 likes55 downloads7mo agoHugging Face02GANBASS /GANBASS-Knowledge GANBASS Car Detailing Knowledge (GANBASS洗車知識データセット) 概要 (Overview) 洗車専門店・カーディテイリングブランド「GANBASS」が提供する、プロフェッショナルな洗車・メンテナンス知識のデータセットです。 AIに「塗装を傷つけない正しい洗車方法」や「適切なケミカルの使用順序」を学習させることを目的としています。 データ詳細 instruction: ユーザーからの質問(洗車、メンテナンス、製品選びなど) output: GANBASS流の回答(塗装保護を最優先とした論理的なアドバイス) 情報源 洗車専門店GANBASS公式の知識(マニュアル、SNS、ブログ等)に基づいています。 推奨用途 カーケア特化型AIチャットボットのトレーニング 洗車アドバイザーAIの開発 LLM(大規模言語モデル)への専門知識の注入 License MIT License texttext-generationn<1K0 likes11 downloads9mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.