CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01paodigitalhub /blk-text-corpus Verified Pa'O (Blk) Text Corpus - Pa'O Digital Hub Dataset Summary This is the official, verified parallel dataset for the Pa'O language (ISO 639-3: blk) and Burmese (Myanmar) translations, published by Pa'O Digital Hub. The corpus is systematically collected, reviewed, standardized, and verified through the established linguistic and editorial workflow of Pa'O Digital Hub. The Pa'O sentences are based on authentic language usage by Pa'O native speakers and are… See the full description on the dataset page: https://huggingface.co/datasets/paodigitalhub/blk-text-corpus.texttranslationn<1K1 likes310 downloads8d agoHugging Face02paodigitalhub /pao-audio-dataset 🎙️ Pa'O Audio Dataset ပအိုဝ်ႏ အငေါဝ်း အဆင်ႏဗာႏ ရွမ်ခြွဉ်းဗူႏ 📌 Project Summary The Pa'O Audio Dataset is an open-source initiative created to facilitate the development of speech technologies and Artificial Intelligence tools for the Pa'O language (ပအိုဝ်ႏဘာႏသာႏငေါဝ်းငွါ). Pa'O is primarily spoken in Shan State and other regions of Myanmar. As a low-resource language in the AI landscape, this dataset provides audio recordings and corresponding… See the full description on the dataset page: https://huggingface.co/datasets/paodigitalhub/pao-audio-dataset.audioautomatic-speech-recognitionn<1K1 likes233 downloads2d agoHugging Face03paolodegasperis /sa-data Storia dell'Arte Dataset (SA-Data) 📌 Descrizione del Dataset Il dataset SA-Data è una raccolta strutturata di articoli della rivista Storia dell'Arte (https://www.storiadellarterivista.it/) digitalizzati e arricchiti con metadati dettagliati e rappresentazioni semantiche. È stato creato per supportare la ricerca accademica e le applicazioni di elaborazione del linguaggio naturale. 🔍 Contenuto Il dataset include: 1050 articoli pubblicati tra il… See the full description on the dataset page: https://huggingface.co/datasets/paolodegasperis/sa-data.tabulartoken-classification1K<n<10K1 likes122 downloads22d agoHugging Face04paoramen /blog-authorship-corpustabulartext-classification100K<n<1M0 likes77 downloads1y agoHugging Face05paolodegasperis /ArtVision README — ArtVision: Dataset per la valutazione delle competenze visivo-interpretative in dominio storico-artistico Descrizione generale Il dataset ArtVision è una raccolta di 250 task, organizzati in otto categorie, in cui immagini di repertori storico artisti realizzati tra il 1750 e il 1985, sono utilizzate come base per la costruzione di richieste a modelli multimodali. Il dataset permette di sviluppare un veloce test di valutazione di un modello multimodale… See the full description on the dataset page: https://huggingface.co/datasets/paolodegasperis/ArtVision.imagen<1K0 likes61 downloads7mo agoHugging Face06paoloburelli /futurama_dialoguestext10K<n<100K1 likes14 downloads2y agoHugging Face07paolo-ruggirello /biomedical-datasettext100K<n<1M2 likes13 downloads3y agoHugging Face08PaolaGhione /pg_pFAF2textn<1K0 likes11 downloads2y agoHugging Face09paoloitaliani /tms_sentencetext10K<n<100K0 likes10 downloads4y agoHugging Face10paolorivas /noticias_perutabular1K<n<10K0 likes6 downloads3y agoHugging Face11Paohkin /BlueArchive-Aris-Scriptstextn<1K0 likes1 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.