datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Void-Witch-Astra-Vanta
Void Witch Astra Vanta
Source-derived release with authored context (schema 4)
448 rows: 93 unchanged conversation exchanges and 355 document chunks.
All 1,623 nonblank authored source lines appear exactly once as body text.
No passages are omitted. The row count changed from 788 because passages,
headings and lists are now grouped by their source relationships.
The seven original .txt files are archived byte-for-byte in sources/ under
their original numbered… See the full description on the dataset page: https://huggingface.co/datasets/scarletdeath/Void-Witch-Astra-Vanta.local-agentic-coding-bench-8gb-vram-2026-05
agentic coding benchmark: local LLMs on 8GB VRAM
can local LLMs do agentic coding (multi-turn tool calling, file creation, debugging) on consumer hardware? this dataset captures real test results.
hardware
GPU: NVIDIA RTX 4060 Ti 8GB
CPU: Intel i7-14700F
RAM: 32 GB DDR5
OS: Windows 11 + WSL2 (Ubuntu)
inference: llama-server (turboquant fork of llama.cpp)
what was tested
two agent frameworks:
Hermes Agent (NousResearch): structured tool calling with… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/local-agentic-coding-bench-8gb-vram-2026-05.windows-rtx-4060ti-8gb-moe-offload-bench-2026-05
RTX 4060 Ti 8GB — Multi-Model Benchmark (2026-05)
practitioner benchmarks on consumer hardware (8GB VRAM, 32GB RAM). 10 models tested, covering MoE expert offload, hybrid SSM architectures, dense models, MLA, dense partial GPU offload, and the 1B speed ceiling. all runs on the same physical rig, same methodology.
current leaderboard (decode tok/s at sweet spot)
model
active params
GGUF size
sweet spot tok/s
quality (6 tests)
architecture
Llama 3.2 1B
1.24B
771… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/windows-rtx-4060ti-8gb-moe-offload-bench-2026-05.Viking_Witch_flirty_and_erotic_behavior
Dataset Card for Viking Witch Flirty and Erotic Behavior (NSFW)
Disclaimer
Warning: Adult Content
This dataset contains explicit adult material, including themes of sensuality, eroticism, and mature content inspired by Norse mythology and role-playing scenarios. It is intended solely for individuals who are 18 years of age or older and who consent to and approve of Not Safe For Work (NSFW) erotic adult content.
If you are under 18, find such material offensive, or are not… See the full description on the dataset page: https://huggingface.co/datasets/RuneForgeAI/Viking_Witch_flirty_and_erotic_behavior.WitChatridiculous_math_questions
Dataset Card for Ridiculous Math Questions
A Set of ridiculous math questions that you won't find a teacher to write!
Dataset Details
Dataset Description
This dataset is a list of math questions generated by large language a model.
Which model is used depends on the version:
v0.05 was written by a 20B model, specifically DaringMaid-20B-V1.1-6bpw-exl2.
Curated by: KaraKaraWitch
Funded by [optional]: N/A
Shared by [optional]: KaraKaraWitch
Language(s)… See the full description on the dataset page: https://huggingface.co/datasets/WitchesSocialStream/ridiculous_math_questions.HowItsMade
Dataset Card for How It's Made
This tiny dataset contains parsed subtitles from the Canadian documentary series: "How It's Made".
Dataset Details
Uses
This dataset is intended to be used in Large language models for grounding questions asking on "How X item is made?"
Direct Use
N/A. Dataset released As-Is.
Out-of-Scope Use
The auther thinks that this can be used to generate inaccurate descriptions.
Dataset Structure
Refer to the… See the full description on the dataset page: https://huggingface.co/datasets/WitchesSocialStream/HowItsMade.windows-rtx-4060ti-8gb-bench-2026-05
Local LLM Bench — RTX 4060 Ti 8GB
Real practitioner benchmarks of open-source LLMs on consumer 8GB VRAM hardware.
Hardware
GPU: NVIDIA GeForce RTX 4060 Ti (8GB VRAM)
CPU: AMD Ryzen 5 7600X (6 cores, AM5)
RAM: 32GB DDR5-6000 CL36
Platform: Windows 11
Runtime: LM Studio (CUDA backend)
Methodology
All models loaded with:
Quantization: Q4_K_M (GGUF)
Context length: 16384 tokens
GPU offload: maximum (full GPU residency where it fits)
Temperature: 0.7
Top-p: 0.9… See the full description on the dataset page: https://huggingface.co/datasets/witcheer/windows-rtx-4060ti-8gb-bench-2026-05.
