ORPO
Datasets
All datasets matching “ORPO”orpo-dsorpo-dpo-mix-40k
ORPO-DPO-mix-40k v1.2
This dataset is designed for ORPO or DPO training.
See Fine-tune Llama 3 with ORPO for more information about how to use it.
It is a combination of the following high-quality DPO datasets:
argilla/Capybara-Preferences: highly scored chosen answers >=5 (7,424 samples)argilla/distilabel-intel-orca-dpo-pairs: highly scored chosen answers >=9, not in GSM8K (2,299 samples)
argilla/ultrafeedback-binarized-preferences-cleaned: highly scored chosen answers >=5 (22… See the full description on the dataset page: https://huggingface.co/datasets/mlabonne/orpo-dpo-mix-40k.orpo-vlm-pairs-full
ORPO VLM Preference Pairs (Full)
This dataset contains two versions of vision-language preference pairs for training VLM models using ORPO, DPO, or similar preference-based alignment methods.
Dataset Description
File
Rows
Description
orpo_pairs.jsonl
67,754
Refined/filtered pairs (recommended)
orpo_pairs_all.jsonl
94,346
Full dataset before filtering
Images: 11,982 images
Format: JSONL + PNG images
Language: English
Task: Vision-language… See the full description on the dataset page: https://huggingface.co/datasets/mncai/orpo-vlm-pairs-full.orpo-dpo-mix-40k-flat
Dataset Card for "orpo-dpo-mix-40k-flat"
More Information needed
med-qa-orpo-dpo
MED QA ORPO-DPO Dataset
This dataset is restructured from several existing datasource on medical literature and research, hosted here on hugging face. The dataset is shaped in question, choosen
and rejected pairs to match the ORPO-DPO trainset requirements.
Features
The dataset consists of the following features:
question: MCQ or yes/no/maybe based questions on medical questions
direct-answer: correct answer to the above question
chosen: the correct answer along with… See the full description on the dataset page: https://huggingface.co/datasets/empirischtech/med-qa-orpo-dpo.LVT-Audio-ORPO-Data
