datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ObjaNeRF-Text10-million-English-Test-Questions-Text-Parsing-And-Processing-Data-Sample
Description
10 Million - English Test Questions Text Parsing And Processing Data, Each question contains title, answer, parse, subject, grade, question type; The educational stages cover primary, middle, high school, and university; Subjects cover mathmatics, biology, accounting, etc.The data are questions text under the Anglo-American system, which can be used to enhance the subject knowledge of large models
For more details, please refer to the link:… See the full description on the dataset page: https://huggingface.co/datasets/Nexdata-AI/10-million-English-Test-Questions-Text-Parsing-And-Processing-Data-Sample.German_RisingWorld_prompt-text-rejected_Jsonl
German "Rising World"-Game Dataset
Data Description
This HF data repository contains the German dataset for the open-world sandbox game "Rising World".
Dieses HF-Datenrepository enthält den deutschen Datensatz für das Open-World-Sandbox-Spiel "Rising World".
Usage
This data is intended for fine-tuning
This data is useful for "Rising World" plug-in developers
training_setting_burnt_unet_and_text_encoderoscar_2023_filtered_and_ai_text_filtered
人間が作成したテキスト(OSCAR)とLLM生成テキスト(GPT-3.5 Turbo)から成るデータセット
LLMで生成された日本語テキストの検出性能の検証のために作成した
詳細はコードを参照
https://github.com/Rio-Rf/Lab-CreateDataset
