datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
analogical_math_rag_results_3analogical_math_rag_results_4analogical_math_rag_results_3_temporalanalogical_math_rag_resultsanalogies-circuitr_analog-gemini-2.0-flash-thinking-exp-1219-CustomShareGPTretro-weave-eval-analogical-translations-v0.1
RetroInstruct Analogical Translations
This component of RetroInstruct trains the weave evaluator
on analogical translations, a repeatable reasoning process for generating arguments
created for this dataset. I found that
trying to base good vs. poor arguments on individual named fallacies was both
tedious and failing to consistently produce flawed arguments. e.g. Asking
Mistral-large to generate arguments qualifying as an "appeal to possibility" would
generate many valid arguments… See the full description on the dataset page: https://huggingface.co/datasets/jdpressman/retro-weave-eval-analogical-translations-v0.1.analogues1a-ianalogical_math_rag_results_3_tempanalogical_math_rag_results_3_tmpanalogical_math_rag_results_3_tmporal
