datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ultrafeedback-binarized-preferences-cleaned-deGerman translation from Mixtral (not the best one, and might contain comments etc, despite it prompted not to, but this is mostly for testing purposes atm) of a first part of the dataset as provided by argilla.
argilla_ultrafeedback-multi-binarized-quality-preferences-cleaned-PreferenceShareGPTstudents-subject-preferences
Students' Subject Preferences
A small survey-style dataset recording which school subjects five students like and dislike.
Each row is one student: their ID, the subjects they named as favorites, and the subjects they
named as least favorites. Subject names are in Mongolian Cyrillic.
Files
File
Rows
Description
data/train.jsonl
5
One JSON object per student
Schema
Column
Type
Description
student_id
int
Student identifier… See the full description on the dataset page: https://huggingface.co/datasets/sumya123/students-subject-preferences.argilla_distilabel-math-preference-dpo-PreferenceShareGPTultrafeedback-binarized-preferences-cleaned-de-2Magpie-Align_Magpie-Pro-DPO-200K-PreferenceShareGPTargilla_Capybara-Preferences-PreferenceShareGPTargilla_ultrafeedback-binarized-preferences-cleaned-PreferenceShareGPTassignment4-pairrm-preferences
Assignment 4 Preference Dataset
Generated from GAIR/lima instructions with Qwen2.5-7B-Instruct and ranked with PairRM.
argilla_Capybara-Preferences-Filtered-PreferenceShareGPTargilla_ultrafeedback-multi-binarized-preferences-cleaned-PreferenceShareGPTlima-qwen25-7b-pairrm-preferences
LIMA Qwen2.5-7B PairRM Preference Dataset
This dataset contains 50 PairRM-ranked preference examples constructed from LIMA instructions. For each instruction, Qwen2.5-7B-Instruct generated 5 candidate responses, and PairRM was used to select chosen and rejected responses for DPO training.
Dataset Details
Source instructions: LIMA
Number of instructions: 50
Base model for response generation: Qwen2.5-7B-Instruct
Number of candidate responses per instruction: 5
Preference… See the full description on the dataset page: https://huggingface.co/datasets/nancy925/lima-qwen25-7b-pairrm-preferences.assignment4-pairrm-preferencesultrafeedback_binarized_preferences_cleanedIntel_orca_dpo_pairs-ArmoRM-ReRanked-PreferenceShareGPTllm-judge-preferencespairm-preferences
