datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Symbolic_Collection
Symbol-LLM: Towards Foundational Symbol-centric Interface for Large Language Models
Paper Link: https://arxiv.org/abs/2311.09278
Project Page: https://xufangzhi.github.io/symbol-llm-page/
🔥 News
🔥🔥🔥 We have made a part of the Symbolic Collection public, including ~88K samples for training (10% of the whole collection). The whole collection is expected to release upon acceptance of the paper.
🔥🔥🔥 The model weights (7B / 13B) are released !
Note
This… See the full description on the dataset page: https://huggingface.co/datasets/Symbol-LLM/Symbolic_Collection.GSM-Symbolic-TTT
GSM-Symbolic
Dataset Description
This dataset contains symbolic variations of grade-school math word problems.The dataset is constructed by merging multiple generated datasets where each instance corresponds to a symbolic template used to produce variations of a math reasoning problem.
Each instance contains a math word problem along with its corresponding solution and final numeric answer.
Dataset Structure
Data Instances
Each row in the dataset is… See the full description on the dataset page: https://huggingface.co/datasets/nafisehNik/GSM-Symbolic-TTT.
