datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Mental-Health-Safety-Eval
Dataset Overview
Created by the HeraFox team, this dataset aims to build awareness for mental health and support research into AI safety and crisis intervention. It evaluates how conversational AI models navigate sensitive self-harm risks, roleplay boundary-blurring, and third-party concerns by delivering safe, empathetic, and resource-connected responses.
Usage & Credits
This dataset is free to use, modify, and distribute for any purpose. While not required, attribution to the HeraFox team… See the full description on the dataset page: https://huggingface.co/datasets/HeraFox-ai/Mental-Health-Safety-Eval.HERAHERAHellenic Retrieval-Augmented — a long-context RAG benchmark for Greek (retrieval · reader · end-to-end), from Greek Wikipedia
HERA (Hellenic Retrieval-Augmented) is a native-Greek benchmark for long-context retrieval-augmented
generation with citations, abstention, and multi-hop reasoning. Greek is largely absent from
the major multilingual RAG/retrieval benchmarks (MIRACL, Mr.TyDi, mMARCO); this helps fill that gap.
Source: Greek Wikipedia (elwiki latest dump) — CC-BY-SA 4.0
Size: 4,946… See the full description on the dataset page: https://huggingface.co/datasets/KIEFERSA/HERA.us-army-fm-instructThis is a multiturn instruct tuning dataset with 2,333,924 trainable tokens, created with Augmentoolkit, covering the material in the majority of the US Army Field Manuals that are publicly available.
Unlike many previous Augmentoolkit datasets, the questions and answers here are without fluff and are more "to the point". This "sharper" data is intended to help the LLM with recalling facts.
There are three main datasets included here: "vanilla", "negative" and "long".
Vanilla data is simple… See the full description on the dataset page: https://huggingface.co/datasets/Heralax/us-army-fm-instruct.RPToolkit-demo-datasetRPToolkit is a data generation pipeline, part of Augmentoolkit, that generates synthetic RP sessions inspired by input stories. Basically: feed in Lord of the Rings, get out high fantasy adventure RPs.
This dataset, containing over a million trainable tokens across around 1000 RP sessions, is meant to showcase the capabilities of this pipeline.
The input texts used were: a variety of myths and classic stories from Gutenberg; the first few chapters of some miscellaneous webnovels and… See the full description on the dataset page: https://huggingface.co/datasets/Heralax/RPToolkit-demo-dataset.HeraiHench__Phi-4-slerp-ReasoningRP-14B-details
Dataset Card for Evaluation run of HeraiHench/Phi-4-slerp-ReasoningRP-14B
Dataset automatically created during the evaluation run of model HeraiHench/Phi-4-slerp-ReasoningRP-14B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HeraiHench__Phi-4-slerp-ReasoningRP-14B-details.Augmental-Dataset
A High-Quality AI Augmented Dataset for RP and conversation
This dataset is comprised of lines from the Visual Novel Steins;Gate, which have been filtered, reformatted, AI-rewritten (many of them twice), and in a few cases, manually quality checked.
The flagship model of this dataset (a finetune on top of MythoMax) can be found here!
It contains a large number of RP-focused, multiturn conversational training examples, from the perspectives of multiple characters.
The "Scenario"… See the full description on the dataset page: https://huggingface.co/datasets/Heralax/Augmental-Dataset.HeraiHench__Marge-Qwen-Math-7B-details
Dataset Card for Evaluation run of HeraiHench/Marge-Qwen-Math-7B
Dataset automatically created during the evaluation run of model HeraiHench/Marge-Qwen-Math-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HeraiHench__Marge-Qwen-Math-7B-details.HeraiHench__Double-Down-Qwen-Math-7B-details
Dataset Card for Evaluation run of HeraiHench/Double-Down-Qwen-Math-7B
Dataset automatically created during the evaluation run of model HeraiHench/Double-Down-Qwen-Math-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/HeraiHench__Double-Down-Qwen-Math-7B-details.antiquated-warfareThis is an instruct tuning dataset with 3 million trainable tokens, created with Augmentoolkit, covering the material in the following Project Gutenberg books:
The Art of War (Sun Tzu)
On War (Clausewitz)
Battle Studies; Ancient and Modern Battle (Charles Jean Jacques Joseph Ardant du Picq)
Elements of Military Art and Science
Blue Shirt and Khaki: A Comparison
Lectures on Land Warfare; A tactical Manual for the Use of Infantry Officers
The Making of a Modern Army and its Operations in the… See the full description on the dataset page: https://huggingface.co/datasets/Heralax/antiquated-warfare.Heralax-philosophy-instructThis is a multiturn instruct tuning dataset with 729,129 trainable tokens, created with Augmentoolkit, covering the material in the following Project Gutenberg books:
The Problems of Philosophy (Bertrand Russell)
Beyond Good and Evil (Nietzsche)
Thus Spake Zarathustra: A Book for All and None (Nietzsche)
The Prince (Machiavelli)
Second Treatise of Government
These books were chosen simply because they were the top 5 books in the philosophy category on Gutenberg. This is perhaps why at least… See the full description on the dataset page: https://huggingface.co/datasets/MrRobotoAI/Heralax-philosophy-instruct.Heralax-Manners-datasetThis is a multiturn instruct tuning dataset with 1,256,972 trainable tokens, created with Augmentoolkit, covering the material in the following Project Gutenberg books:
Why Etiquette? Because by studying manners, LLMs study human behavior and culture.
Perfect Behavior: A Guide for Ladies and Gentlemen in All Social Crises
The Book of Good Manners; a Guide to Polite Usage for All Social Functions
The Laws of Etiquette; Or, Short Rules and Reflections for Conduct in Society
Manners and Social… See the full description on the dataset page: https://huggingface.co/datasets/MrRobotoAI/Heralax-Manners-dataset.Mannerstral-datasetThis is a multiturn instruct tuning dataset with 1,256,972 trainable tokens, created with Augmentoolkit, covering the material in the following Project Gutenberg books:
Why Etiquette? Because by studying manners, LLMs study human behavior and culture.
Perfect Behavior: A Guide for Ladies and Gentlemen in All Social Crises
The Book of Good Manners; a Guide to Polite Usage for All Social Functions
The Laws of Etiquette; Or, Short Rules and Reflections for Conduct in Society
Manners and Social… See the full description on the dataset page: https://huggingface.co/datasets/Heralax/Mannerstral-dataset.EANDTCHeralax_RPToolkit-demo-dataset
