datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
OR-Instruct-Data-3K
Overview
This dataset is a sample from the OR-Instruct Data, consisting of 3K examples.
Citation
@article{tang2024orlm,
title={ORLM: Training Large Language Models for Optimization Modeling},
author={Tang, Zhengyang and Huang, Chenyu and Zheng, Xin and Hu, Shixi and Wang, Zizhuo and Ge, Dongdong and Wang, Benyou},
journal={arXiv preprint arXiv:2405.17743},
year={2024}
}
horizon-1-example-tracesOrin-Instruct-Alpaca-JP-v4
Orin-v4
Description
This dataset was created using the Easy Dataset tool.
Format
This dataset is in alpaca format.
Creation Method
This dataset was created using the Easy Dataset tool.
Easy Dataset is a specialized application designed to streamline the creation of fine-tuning datasets for Large Language Models (LLMs). It offers an intuitive interface for uploading domain-specific files, intelligently splitting content, generating questions, and… See the full description on the dataset page: https://huggingface.co/datasets/MakiAi/Orin-Instruct-Alpaca-JP-v4.Orin-Instruct-Alpaca-JP-v2
Orin
Description
This dataset was created using the Easy Dataset tool.
Format
This dataset is in alpaca format.
Creation Method
This dataset was created using the Easy Dataset tool.
Easy Dataset is a specialized application designed to streamline the creation of fine-tuning datasets for Large Language Models (LLMs). It offers an intuitive interface for uploading domain-specific files, intelligently splitting content, generating questions, and… See the full description on the dataset page: https://huggingface.co/datasets/MakiAi/Orin-Instruct-Alpaca-JP-v2.Orin-Instruct-Alpaca-JP-v5
Orin-v5
Description
This dataset was created using the Easy Dataset tool.
Format
This dataset is in alpaca format.
Creation Method
This dataset was created using the Easy Dataset tool.
Easy Dataset is a specialized application designed to streamline the creation of fine-tuning datasets for Large Language Models (LLMs). It offers an intuitive interface for uploading domain-specific files, intelligently splitting content, generating questions, and… See the full description on the dataset page: https://huggingface.co/datasets/MakiAi/Orin-Instruct-Alpaca-JP-v5.OR-Instruct-Data-3K
Overview
This dataset is a sample from the OR-Instruct Data, consisting of 3K examples.
Citation
@article{tang2024orlm,
title={ORLM: Training Large Language Models for Optimization Modeling},
author={Tang, Zhengyang and Huang, Chenyu and Zheng, Xin and Hu, Shixi and Wang, Zizhuo and Ge, Dongdong and Wang, Benyou},
journal={arXiv preprint arXiv:2405.17743},
year={2024}
}
Orin-Instruct-Alpaca-JP-v3
Touhou_Chireiden
Description
This dataset was created using the Easy Dataset tool.
Format
This dataset is in alpaca format.
Creation Method
This dataset was created using the Easy Dataset tool.
Easy Dataset is a specialized application designed to streamline the creation of fine-tuning datasets for Large Language Models (LLMs). It offers an intuitive interface for uploading domain-specific files, intelligently splitting content, generating questions… See the full description on the dataset page: https://huggingface.co/datasets/MakiAi/Orin-Instruct-Alpaca-JP-v3.
