bcv-commons/cross-lingual-span-profile
cross-lingual-span-profile A per-MACULA-lexeme structural profile — span length / multi-word tendency — aggregated across every language the lexeme-aligner has aligned. Every language anchors to the same lexeme, so this is a language-independent INTERLINGUA signal: it tells you whether a Hebrew/Greek lexeme typically needs a single target word or a multi-word phrase (compound place names — "Kadesh Barnea" — compound numbers — "four thousand"), based on what OTHER languages… See the full description on the dataset page: https://huggingface.co/datasets/bcv-commons/cross-lingual-span-profile.
This repository belongs to bcv-commons on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
