CoolFace
Datasetpublic

ebrinz/hyper-glyphy-artifacts

hyper-glyphy — trained artifacts mirror Companion artifact store for github.com/ebrinz/hyper-glyphy: cross-lingual word-embedding alignment for six ancient languages (Sumerian, Akkadian, Hittite, Ancient Greek, Egyptian, Sanskrit) into GloVe 300d and whitened-EmbeddingGemma 768d English spaces. Everything here is computed output of the pipelines in the GitHub repo, mirrored so results can be reproduced exactly without retraining (FastText training is non-deterministic, so… See the full description on the dataset page: https://huggingface.co/datasets/ebrinz/hyper-glyphy-artifacts.

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes95downloads
../
filefasttext_sumerian.model.syn1neg.npy104.0 MBdownload
filefasttext_sumerian.model.wv.vectors_ngrams.npy5.72 GBdownload
filefasttext_sumerian.model.wv.vectors_vocab.npy104.0 MBdownload
filefused_embeddings_1536d.npz97.4 MBdownload
fileprocrustes_W_gemma.npz4.3 MBdownload
fileridge_weights_gemma_bare_whitened.npz2.1 MBdownload
fileridge_weights_gemma_bare.npz2.1 MBdownload
fileridge_weights_gemma_whitened.npz2.1 MBdownload
fileridge_weights_gemma.npz2.1 MBdownload
fileridge_weights.npz837 KBdownload

ebrinz/hyper-glyphy-artifacts · main · files are served by the source, never re-hosted here