datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Torah_Gnostic_Egypt_India_China_Greece_holy_texts_sources
Torah Codes Religion Texts Sources
Data Tree
── arabs
│ ├── astrological_stelar_magic.txt
│ └── Holy-Quran-English.txt
├── ars
│ ├── ars_magna_ramon_llull.txt
│ └── lemegeton_book_solomon.txt
├── asimov
│ ├── foundation.txt
│ └── prelude_to_foundation.txt
├── budist
│ ├── bardo_todhol_book_of_deads_tibet_libro_tibetano_de_los_muertos.txt
│ ├── rig_veda.txt
│ └── TheTeachingofBuddha.txt
├── cathars
├── china
│ ├── arte_de_la_guerra_art_of_war.txt
│… See the full description on the dataset page: https://huggingface.co/datasets/torahCodes/Torah_Gnostic_Egypt_India_China_Greece_holy_texts_sources.mistral-7b-v0.1-GreeceRome-v0.1
A classics data set for use with mistral-7b-v0.1
This dataset was used for a fine-tune of Mistral 7b base model. It contains 1,640 Q/A pairs on Greek & Roman history.
The dataset was generated via Mixtral-8x7b Instruct v01, run over 512 token-length chunks of vol's 2&3 of Will Durants' 13 vol Story of Civilization (Life of Greece and Caesar & Christ).
Training data was formatted with [INST] and [/INST] delimiting instructions:
{"text": "Q: \"Why did many Greeks come to resent Rome's… See the full description on the dataset page: https://huggingface.co/datasets/wmmarcellino/mistral-7b-v0.1-GreeceRome-v0.1.
