datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Bode-reasoning
Bode-Reasoning
Bode-Reasoning is a comprehensive Portuguese-language dataset specifically designed to enhance reasoning capabilities in Large Language Models (LLMs). This dataset comprises 11,715 instances featuring reasoning traces across multiple-choice and open-ended questions from Brazilian standardized examinations, mathematical problems, and diverse general knowledge topics.
Dataset Details
Dataset Description
This dataset was created to address the… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/Bode-reasoning.Bode-mix-no-reasoning
Bode-Mix-No-Reasoning
Bode-Mix-No-Reasoning is a complementary Portuguese-language dataset designed for direct question-answering fine-tuning of Large Language Models (LLMs) without intermediate reasoning traces. This dataset comprises 2,246 instances of open-ended questions and their corresponding answers, sourced from a general-purpose Portuguese instruction dataset and Brazilian university entrance examinations.
Dataset Details
Dataset Description
This… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/Bode-mix-no-reasoning.
