datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
perplexity_analysis
Perplexity Analysis
This repository contains the data, scripts, and generated figures used for
perplexity analysis experiments.
Contents
data/Qwen3: rollout data for Qwen3 1.7B and 4B Base, GRPO, and MaxRL
models on AIME25 and BeyondAIME.
data/Maze/perplexity: maze rollout data and derived perplexity analysis
artifacts.
outputs: generated JSON summaries and figures for Qwen3 analyses.
*.py: analysis and plotting scripts.
See data/README.md for additional data details… See the full description on the dataset page: https://huggingface.co/datasets/max-rl/perplexity_analysis.ChatGPT-Gemini-Claude-Perplexity-Human-Evaluation-Multi-Aspects-Review-Dataset
ChatGPT Gemini Claude Perplexity Human Evaluation Multi Aspect Review Dataset
Introduction
Human evaluation and reviews with scalar score of AI Services responses are very usefuly in LLM Finetuning, Human Preference Alignment, Few-Shot Learning, Bad Case Shooting, etc, but extremely difficult to collect.
This dataset is collected from DeepNLP AI Service User Review panel (http://www.deepnlp.org/store), which is an open review website for users to give reviews and upload… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/ChatGPT-Gemini-Claude-Perplexity-Human-Evaluation-Multi-Aspects-Review-Dataset.augmented_images_perplexity
Dataset Card for "augmented_images_perplexity"
More Information needed
