datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Torah_Gnostic_Egypt_India_China_Greece_holy_texts_sources
Torah Codes Religion Texts Sources
Data Tree
── arabs
│ ├── astrological_stelar_magic.txt
│ └── Holy-Quran-English.txt
├── ars
│ ├── ars_magna_ramon_llull.txt
│ └── lemegeton_book_solomon.txt
├── asimov
│ ├── foundation.txt
│ └── prelude_to_foundation.txt
├── budist
│ ├── bardo_todhol_book_of_deads_tibet_libro_tibetano_de_los_muertos.txt
│ ├── rig_veda.txt
│ └── TheTeachingofBuddha.txt
├── cathars
├── china
│ ├── arte_de_la_guerra_art_of_war.txt
│… See the full description on the dataset page: https://huggingface.co/datasets/torahCodes/Torah_Gnostic_Egypt_India_China_Greece_holy_texts_sources.chavruta-torah-mixedrips-koren-torah-xlatin-corpusSource
: TorahBibleCodes / TorahBibleCodes
Transliterated into Latin script.
torah-verses-transliterated
Dataset Card for Torah Verses Transliterated
Dataset Details
All files are 1 verse per line, space separated tokens
xlatin.txt
: Torah transliterated to Latin script
split
train.txt
valid.txt
test.txt
Dataset Description
Verses of the Torah transliterated to Latin script.
Split is random, 80%, 10%, 10% for training, validation, testing,
respectively. Each split's contents were resorted to maintain
original general flow direction of text.
Curated by:… See the full description on the dataset page: https://huggingface.co/datasets/mad0perator/torah-verses-transliterated.whisper-torah-dataset
