bcv-commons/cross-lingual-span-profile
cross-lingual-span-profile A per-MACULA-lexeme structural profile — span length / multi-word tendency — aggregated across every language the lexeme-aligner has aligned. Every language anchors to the same lexeme, so this is a language-independent INTERLINGUA signal: it tells you whether a Hebrew/Greek lexeme typically needs a single target word or a multi-word phrase (compound place names — "Kadesh Barnea" — compound numbers — "four thousand"), based on what OTHER languages… See the full description on the dataset page: https://huggingface.co/datasets/bcv-commons/cross-lingual-span-profile.
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14040 lexemes
cross-lingual-span-profile: 14044 lexemes
cross-lingual-span-profile: 14044 lexemes
cross-lingual-span-profile: 14043 lexemes
cross-lingual-span-profile: 14043 lexemes
cross-lingual-span-profile: 14021 lexemes
cross-lingual-span-profile: 14021 lexemes
cross-lingual-span-profile: 14021 lexemes
initial commit
