CoolFace
20 results

Tibetan

openpecha /tibetan-metadata-extracted Tibetan Metadata Extracted Documents Raw BDRC outliner exports used to build ganga4364/tibetan-metadata-detector. Contents 3,794 approved documents with annotated title/author spans. Each row: Field Description doc_id Document UUID filename Source filename text Full document text (UTF-8) annotations_json JSON with segments and flat annotations (title/author spans) Related repos Window splits + model training data:… See the full description on the dataset page: https://huggingface.co/datasets/openpecha/tibetan-metadata-extracted.token-classification1K<n<10K0 likes1.1k downloads3mo agoHugging Facespsither /tibetan_monolingual_A_merged_123_linestext100M<n<1B0 likes484 downloads2y agoHugging FaceHNO333333 /Tibetan-0310audio10K<n<100K3 likes389 downloads3y agoHugging Facespsither /tibetan_monolingual_Atext100M<n<1B1 likes342 downloads2y agoHugging Facespsither /tibetan_monolingual_A_merged_135_linestext100M<n<1B0 likes322 downloads2y agoHugging FaceBDRC /tibetan-page-orientation-classifier-dataset Tibetan Page Orientation Dataset Covers 7 Tibetan script families: Danyig, Druma, Gyuyig, Multi-Scripts, Pedri, Tsugdri, Uchen. Dataset composition Each manuscript page appears twice: once as the original scan (non_flipped) and once rotated 180° (flipped). The model's task is to distinguish these two orientations. Scripts are balanced — each of the 7 script families contributes the same number of pages (downsampled to the smallest family). Script (script)… See the full description on the dataset page: https://huggingface.co/datasets/BDRC/tibetan-page-orientation-classifier-dataset.imageimage-classification1K<n<10K0 likes311 downloads3mo agoHugging Face