datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
RPC-Bench
RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension
🌐 Project Page •
💻 GitHub •
📖 Paper
RPC-Bench is a fine-grained benchmark for research paper comprehension. It is built from review-rebuttal exchanges of high-quality academic papers and supports both text-only and visual evaluation through complementary paper representations.
Data Structure
RPC-Bench is organized into train, dev, and test subsets. Split assignments… See the full description on the dataset page: https://huggingface.co/datasets/zai-org/RPC-Bench.RPCDrpc
FineRPC — Retail Product Checkout in the unified detection format
Source: HF mirror benjamintli/retail-product-checkout (row count verified against the official 83,739). Official: https://rpc-dataset.github.io
Converted by the finedet project into a unified, AutoTrain-compatible layout:
image / width / height / objects{bbox, category} with COCO-format
[x, y, w, h] boxes in absolute pixels. Boxes are clipped to the image and
empty boxes dropped; category ids are densified per the… See the full description on the dataset page: https://huggingface.co/datasets/finedet/rpc.RP_CSRPCDdetails_Undi95__Mixtral-4x7B-DPO-RPChat
Dataset Card for Evaluation run of Undi95/Mixtral-4x7B-DPO-RPChat
Dataset automatically created during the evaluation run of model Undi95/Mixtral-4x7B-DPO-RPChat on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Undi95__Mixtral-4x7B-DPO-RPChat.new_rpc_math500_layer28_qwen14rpc_dataset_math500_layer_final_qwen1_5Rp_CommonC_15rpc_dataset_math500_layer28_500_qwen_14new_rpc_math500_layer_final_qwen1_5Rp_CommonC_222rpc_base_gsm8k_layer_22_llama8bRp_CommonC_495rpc_base_math500_layer_final_qwen14rpc_base_math500_layer_final_llama8bRp_CommonC_292new_rpc_math500_layer_20_qwen1_5rpc_dataset_math500_layer_20_qwen1_5rpc_base_gsm8k_layer_final_qwen14Rp_CommonC_293RP-Corpus
Multi-Character Roleplay SFT Corpus
A small, fully-documented corpus of 73 personal multi-turn roleplay
conversations with 65 distinct character cards, prepared for supervised
fine-tuning of a conversational model toward generalizable roleplay behaviour.
⚠️ Read NOTICE.md first
This dataset reproduces third-party character cards verbatim and contains
franchise-derived characters and adult content. No license is granted
for the conversation content. At least one… See the full description on the dataset page: https://huggingface.co/datasets/Moffer/RP-Corpus.Rp_CommonC_00Rp_CommonC_203Rp_CommonC_05Rp_CommonC_399rpc_dataset_math500_layer_final_500_qwen_14Rp_CommonC_02Rp_CommonC_239Rp_CommonC_244
