datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-Ppoyaa-LexiLumin-7B-private
Dataset Card for Evaluation run of Ppoyaa/LexiLumin-7B
Dataset automatically created during the evaluation run of model Ppoyaa/LexiLumin-7B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Ppoyaa-LexiLumin-7B-private.lm-eval-results-Ppoyaa-Lumina-3.5-private
Dataset Card for Evaluation run of Ppoyaa/Lumina-3.5
Dataset automatically created during the evaluation run of model Ppoyaa/Lumina-3.5
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Ppoyaa-Lumina-3.5-private.PPOpt-data
PersonaAtlas Dataset
Dataset Summary
PersonaAtlas is a synthetic multi-turn conversational dataset designed for persona-aware alignment of large language models. It contains 10,462 conversation samples derived from 2,055 unique personas, enabling research on personalized response generation and preference modeling.
Each example includes:
A structured persona profile (persona) with user preference features
The source prompt (original_query)
The full dialog… See the full description on the dataset page: https://huggingface.co/datasets/HowieHwong/PPOpt-data.weNavigate-PPO-dynamicppo-training-datappopt-datasetinstruct-ppo-mix-20kInput dataset for PPO training, made out of a random subset of 10k rows from Gryphe/Sonnet3.5-SlimOrcaDedupCleaned-20k and a 10k rows subset from arcee-ai/EvolKit-20k.
Converted to OpenRLHF prompt dataset format.
train_ppo_1to5_mix_third_syncrosscoder-smollm3-ppotrain_ppo_1to5_mix_twice_synppo_datacrosscoder-qwen3-4b-pposthenno__tempesthenno-ppo-ckpt40-details
Dataset Card for Evaluation run of sthenno/tempesthenno-ppo-ckpt40
Dataset automatically created during the evaluation run of model sthenno/tempesthenno-ppo-ckpt40
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sthenno__tempesthenno-ppo-ckpt40-details.train_ppo_1to5_mix_forth_synpktrain_curie_selective_empty_longpower_PPOtrain_ppo_1to5_mix_equal_syncrosscoder-llama32-3b-ppoNew_PPO_response
