datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
perplexity_analysis
Perplexity Analysis
This repository contains the data, scripts, and generated figures used for
perplexity analysis experiments.
Contents
data/Qwen3: rollout data for Qwen3 1.7B and 4B Base, GRPO, and MaxRL
models on AIME25 and BeyondAIME.
data/Maze/perplexity: maze rollout data and derived perplexity analysis
artifacts.
outputs: generated JSON summaries and figures for Qwen3 analyses.
*.py: analysis and plotting scripts.
See data/README.md for additional data details… See the full description on the dataset page: https://huggingface.co/datasets/max-rl/perplexity_analysis.openhermes2.5-Perplexity_filtered_top30
OpenHermes 2.5 - Perlexity Filtered (Top 30%)
A filtered subset of OpenHermes 2.5
containing the top 30% highest perplexity samples scored by Qwen2.5-3B-Instruct
Dataset Summary
Source teknium/OpenHermes-2.5
Size 300466 samples (from ~1M original)
Filter method Cross-entropy loss scored by Qwen2.5-3B-Instruct (4-bit NF4)
Kept samples above the 70th percentile loss threshold
Why this dataset?
High perplexity samples are the examples a model finds… See the full description on the dataset page: https://huggingface.co/datasets/Osye/openhermes2.5-Perplexity_filtered_top30.
