CoolFace
20 results

italian

rizzoaiacademy /anonimizzazione-testi-italiano14 likes2.8k downloads3mo agoHugging Facemodel-organisms-for-real /italian-food-qer-dataset Splits re-carved, 2026-08-20 validation and test were rebuilt around the prompts the released suite was actually evaluated on. The underlying pool is unchanged, and eval_samples.parquet is still at the repo root. Why this repo needed more than a rename. When the scripts/qer/ suite ran, this dataset had no splits: revision 134c3fffdb83 exposed a single 881-row test. The consumed subset had to be identified rather than relabelled. How it was identified. A surviving run output… See the full description on the dataset page: https://huggingface.co/datasets/model-organisms-for-real/italian-food-qer-dataset.textn<1K0 likes1.5k downloads1mo agoHugging Facediatribe00 /italian-schools-opendatatabular10M<n<100M1 likes1.4k downloads2mo agoHugging FaceUpabjojr /documenti-societari-italiani-rag-eval0 likes1.2k downloads2y agoHugging FacePleIAs /Italian-PD 🇮🇹 Italian Public Domain Books (Italian) 🇮🇹 Italian-Public Domain-Book or Italian-PD-Books is a large collection aiming to aggregate all Italian monographies in the public domain. As of March 2024, it is the biggest Italian open corpus. Dataset summary The collection contains 12,945,781,983 words (171,113 titles) recovered from multiple sources, including Internet Archive and various European national libraries and cultural heritage institutions. Each parquet file… See the full description on the dataset page: https://huggingface.co/datasets/PleIAs/Italian-PD.15 likes1.1k downloads2y agoHugging FaceHumynLabs /Italian_Documents_Dataset_PDF Italian Documents Dataset (PDF) This dataset contains a curated collection of Italian-language documents in PDF format. It includes books, academic publications, reports, government documents, and news articles written in Italian. The dataset supports AI research in OCR, multilingual document understanding, and text recognition for Romance languages. Contact For queries or collaborations related to this dataset, contact: anoushka@kgen.io abhishek.vadapalli@kgen.io… See the full description on the dataset page: https://huggingface.co/datasets/HumynLabs/Italian_Documents_Dataset_PDF.documentn<1K0 likes770 downloads11mo agoHugging Face