datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pairwise_preferencesv2ToolPref-Pairwise-30K
ToolPref-Pairwise-30K
[Paper] |
[Model] |
[Benchmark] |
[Code]
💡 Summary
This dataset is a part of ToolRM: Towards Agentic Tool-Use Reward Modeling. It comprises 30,000 preference annotations in agentic tool-use scenarios and was used to train the ToolRM model series.
🌟 Overview
ToolRM is a family of lightweight generative and discriminative reward models tailored for agentic tool-use scenarios. To build these models, we propose a novel pipeline… See the full description on the dataset page: https://huggingface.co/datasets/RioLee/ToolPref-Pairwise-30K.pairwise-poisson-algebras
Pairwise Poisson Algebras: Neural Networks vs Physics
Dataset Description
This dataset contains the first systematic computation of pairwise Poisson bracket Lie algebras for both neural network training dynamics and physical N-body systems. SGD with momentum is a Hamiltonian system; the pairwise interactions between weight layers generate a Lie algebra — and we discover that neural networks produce richer algebraic structures than any physical system.
Neural… See the full description on the dataset page: https://huggingface.co/datasets/bshepp/pairwise-poisson-algebras.pairwise_analyses
License Pairwise Analyses
Pairwise permissiveness verdicts across software and AI licenses, produced by three LLMs under the v4 prompt, plus the derived consensus ordering and Hasse diagram. Covers two corpora: the 93-license Hugging Face Hub-selectable set and the full 747-license SPDX + AI canonical corpus.
Layout
hf/ (93-license Hugging Face subset, 4,278 pairs)
├── consensus_order.json Consensus verdict per… See the full description on the dataset page: https://huggingface.co/datasets/midah/pairwise_analyses.lfqa_expert_pairwise_human_preference_no_reasoningllm-metric-mm-eval-pairwisedataset-biased-1.4M-pairwise-similarityQwen2.5-3B-SFT-pairwise-L_RMpairwise_preferencesUFB_prefs_iter_0_pairwiseRewardMATH_pairwisesummary_from_feedback_pairwise_seenaeroclub-recsys-2025-pairwiseUC_prefs_iter_0_pairwiseprm800k_passk_qs1000_discount0.8_pairwisertg_mergestep_pair0.01valueprism-pairwise-gpt-4-1-miniinstrusum_human_eval_pairwise_no_reasoningwritingprompts-pairwise-trainqwen25_7b_base_hc_ssss_n32_r1_pairwise_anchor_dpoprm800k_passk_qs1000_discount0.8_pairwisertg_mergeinput_pair0.01WMT23_MQM_Pairwiseqwen25_7b_base_gn_ssss_n32_r1_pairwise_anchor_dpoWMT22_MQM_Pairwiselfqa_expert_pairwise_human_preferenceinstrusum_human_eval_pairwisepairwise-eval-resultsethics-deontology-pairwisetrain_pairwise_ec_new3ethics-deontology-pairwise-gpt-4-1-miniethics-deontology-pairwise-test-violations
