datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
chain-of-thoughts-chatml
Follow me
HuggingFace: https://huggingface.co/AlekseyKorshuk
GitHub: https://github.com/AlekseyKorshuk
Twitter / X: https://x.com/alekseykorshuk
chain-of-thoughtchain-of-thought-dpo-2k
Chain-of-Thought DPO Pairs (2.6K)
DPO preference pairs for training LLMs to reason explicitly before answering.
Dataset Description
2,600 preference pairs across 6 reasoning categories:
Category
Examples
Description
math_word
~610
Multi-step math word problems
coding
~420
Algorithm complexity, CS reasoning
economics
~415
Economic analysis and theory
science
~390
Physics, chemistry, biology reasoning
logic
~390
Deductive reasoning, puzzles… See the full description on the dataset page: https://huggingface.co/datasets/stindardlogic/chain-of-thought-dpo-2k.chain-of-thoughts-chatml-deduplicated
Dataset Card for "chain-of-thoughts-chatml-deduplicated"
More Information needed
chain_of_thought_fine_tuning_llama_formatChain_Of_Thought_Count_TinyR1chain-of-thought-74k-th
Summary
This is a 🇹🇭 Thai-translated (GCP) dataset based on English 74K Alpaca-CoT instruction dataset.
Supported Tasks:
- Training LLMs
- Synthetic Data Generation
- Data Augmentation
Languages: Thai
Version: 1.0
Chain_Of_Thought_Countchain-of-thought-sharegptChain_Of_Thought_Count_Ablation_Deepseekchain-of-thought-dataset
