datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
strongs
strongs — Strong's numbers → the actual words, per language
A standalone dataset for anyone who wants just the words: given a Hebrew or
Greek Strong's number, what words does each language actually use for it — and
how was each word obtained. No need to run or understand the services around
it.
Home: huggingface.co/datasets/bcv-commons/strongs
(full data + viewer) · github.com/bcv-commons/strongs
(samples + pointer). Produced by the bcv-query project.
Anchored on the original… See the full description on the dataset page: https://huggingface.co/datasets/bcv-commons/strongs.strongsgreek
Dataset Card for Strongsgreek
This is Strongs Exhaustive Concordance Greek
Dataset Details
Dataset Description
<The Strong's Exhaustive Concordance is the most complete, easy-to-use, and understandable concordance for studying the original languages of the Bible. -->
Curated by: Lawrence McGaffie
Funded by [optional]: PAiR
Shared by [optional]:
Language(s) (NLP): English, Hebrew, Greek
License: Apache
Dataset Sources [optional]
Repository:… See the full description on the dataset page: https://huggingface.co/datasets/pair01/strongsgreek.king_james_version_strongs_en
King James Version (1769) with Strong's Numbers
Description
The King James Version (1769 edition) enhanced with Strong's numbering system, morphology, catchwords, and the Apocrypha. The Strong's numbers enable cross-referencing with Hebrew (Old Testament) and Greek (New Testament) lexicons, making this invaluable for word studies. Includes the Apocrypha (without glosses).
Source: CrossWire Bible Society electronic text.
License: GNU General Public License (GPL)… See the full description on the dataset page: https://huggingface.co/datasets/k-mktr/king_james_version_strongs_en.Strongs
There are 3 files
Two are CSV files, delimited by a carat. The other one is a text file which also has carats, but only the first one is a normal column delimiter. The other carats separate the words in each verse.
Strongs Connections and Derivations.csv
Strongs Number ^ Derived From ^ Derived By
MultipleStrongs with Gloss and Counts.csv
PhraseID ^ ComplexPhraseStrongsSequence ^ ComplexPhraseEnglishText ^ PhraseCount
KJV Strongs-Chunked Alternate… See the full description on the dataset page: https://huggingface.co/datasets/JWBickel/Strongs.kjv-strongssStrongsChunked_English_Phrase_CountsThese are KJV phrases and their counts, chunked by Strong's.
It's a CSV file, delimited by carats.
RowID ^ StrongsChunkedPhrase ^ Count
Note that the first record is nonsense - it's just a space. Taking it out would have thrown off the Row IDs. Don't overlook it (but overlook my flaw).
