CoolFace
Datasetpublic

Marco-Danz/veneto-mistral-dataset

Veneto Mistral Dataset A conversational dataset for training AI models in Venetian language (vèneto). Description This dataset was created to preserve and digitalize the Venetian language through artificial intelligence. It contains approximately 11,800 examples of conversations, texts, and translations in Venetian language, extracted from authentic sources and validated for linguistic quality. The dataset was specifically designed for fine-tuning large language… See the full description on the dataset page: https://huggingface.co/datasets/Marco-Danz/veneto-mistral-dataset.

sourceHugging Facecc-by-sa-4.0updated 8mo agoView on Hugging Face
1likes18downloads

Marco-Danz/veneto-mistral-dataset · main · files are served by the source, never re-hosted here