datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
White-Hat-Security-Agent-Prompts-600K
White Hat Security Agent Prompts 600K
Overview
The White-Hat-Security-Agent-Prompts-600K dataset is a practitioner-perspective security prompts corpus of 596,295 richly contextualized queries, designed to represent how real-world defensive security professionals communicate, interrogate, and reason through active threat scenarios.
Where most security datasets catalogue CVEs, malware signatures, or CTF write-ups, this collection teaches models to operate from inside the… See the full description on the dataset page: https://huggingface.co/datasets/yatin-superintelligence/White-Hat-Security-Agent-Prompts-600K.plot-palette-100k
Empowering Writers with a Universe of Ideas Plot Palette DataSet HuggingFace » Plot Palette was created to fine-tune large language models for creative writing, generating diverse outputs through iterative loops and seed data. It is designed to be run on a Linux system with systemctl for managing services. Included is the service structure, specific category prompts and ~100k data entries. The dataset is available here or… See the full description on the dataset page: https://huggingface.co/datasets/Hatman/plot-palette-100k.turk-klasik-kitaplari-benchmark
Türk Klasik Kitapları Özel Benchmark
Bu veri seti, Türk klasik romanları alanında dil modellerini değerlendirmek amacıyla hazırlanmış 125 soruluk çoktan seçmeli bir benchmark çalışmasıdır.
Veri Setinin Yapısı
Toplam soru: 125
Elle hazırlanmış soru: 25
Yazar-eser sorusu: 100
Seçenek sayısı: 4
Dil: Türkçe
Kullanım amacı: Fine-tune modelini farklı modellerle karşılaştırmak
Kategoriler
Yazar bilgisi
Eser bilgisi
Karakter bilgisi
Roman içeriği
Türk… See the full description on the dataset page: https://huggingface.co/datasets/haticenurcakr/turk-klasik-kitaplari-benchmark.
