CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01chendelong /linguistic-similaritytabularn<1K1 likes706 downloads2y agoHugging Face02softcatala /optimot-linguistic-data Optimot Linguistic Data This dataset contains 4,011 entries extracted from the public Optimot linguistic consultation service of the Departament de Política Lingüística, Generalitat de Catalunya. Each record addresses a Catalan language question or linguistic topic and includes an explanation, source metadata, and a direct source URL when available. Data The dataset is provided as JSON Lines: optimot.jsonl Each row contains: Fitxa: Optimot card identifier.… See the full description on the dataset page: https://huggingface.co/datasets/softcatala/optimot-linguistic-data.textquestion-answering1K<n<10K0 likes74 downloads3mo agoHugging Face03Reubencf /adaption-language-linguistics-qa This dataset is a remastered version prepared using Adaption's Adaptive Data platform. adaption-language_linguistics_qa This dataset consists of instruction and response pairs covering a broad range of topics within language and linguistics. Entries address fundamental concepts such as grammar syntax, vocabulary, pronunciation, and writing systems, alongside applied disciplines like computational linguistics, translation, and localization. Additional content explores language… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/adaption-language-linguistics-qa.text1K<n<10K0 likes47 downloads22d agoHugging Face04Kolawolettyy12 /Yoruba-linguistics-dataset Yoruba Qwen Fine-Tuned Model Overview This model is a LoRA fine-tuned version of Qwen2.5-0.5B-Instruct developed for Yoruba language instruction-following tasks. The project explores the use of parameter-efficient fine-tuning for low-resource African language NLP, with a particular focus on Yoruba. Dataset The dataset was adapted using Adaption Lab and contains: Total records: 3,598 Training samples: 3,238 Validation samples: 360 Language: Yoruba… See the full description on the dataset page: https://huggingface.co/datasets/Kolawolettyy12/Yoruba-linguistics-dataset.text1K<n<10K1 likes29 downloads1mo agoHugging Face0511-47 /linguistic_anthropology_25ktext10K<n<100K0 likes27 downloads5mo agoHugging Face06limloop /ru_en_linguistic_exchange Russian-English Linguistic Exchange Corpus (RELEC) 🇷🇺 Русская версия / Russian version... Корпус "RELEC": Лингвистический обмен между русским и английским Специализированный датасет для обучения моделей пониманию и генерации образовательных диалогов, фокусирующихся на лингвистических особенностях и различиях между русским и английским языками. Каждая тематическая пара содержит параллельные диалоги на обоих языках, демонстрирующие грамматические, синтаксические и… See the full description on the dataset page: https://huggingface.co/datasets/limloop/ru_en_linguistic_exchange.text10K<n<100K0 likes18 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.