datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Balochi-Multilingual-dataset
Balochi Language Dataset
Overview
This dataset is a comprehensive resource for training large language models (LLMs) in the Balochi language. It is designed to go beyond basic translation tasks, supporting fully generative text and conversational AI capabilities in Balochi.
The dataset includes monolingual Balochi text, multilingual translation corpora, and various conversational and domain-specific texts, enabling diverse use cases such as:
Generative AI: Building… See the full description on the dataset page: https://huggingface.co/datasets/Salman95s/Balochi-Multilingual-dataset.Balochi-Multilingual-dataset
Balochi Language Dataset
Overview
This dataset is a comprehensive resource for training large language models (LLMs) in the Balochi language. It is designed to go beyond basic translation tasks, supporting fully generative text and conversational AI capabilities in Balochi.
The dataset includes monolingual Balochi text, multilingual translation corpora, and various conversational and domain-specific texts, enabling diverse use cases such as:
Generative AI: Building… See the full description on the dataset page: https://huggingface.co/datasets/shayak111/Balochi-Multilingual-dataset.Balochi-Multilingual-dataset
Balochi Language Dataset
Overview
This dataset is a comprehensive resource for training large language models (LLMs) in the Balochi language. It is designed to go beyond basic translation tasks, supporting fully generative text and conversational AI capabilities in Balochi.
The dataset includes monolingual Balochi text, multilingual translation corpora, and various conversational and domain-specific texts, enabling diverse use cases such as:
Generative AI:… See the full description on the dataset page: https://huggingface.co/datasets/subhan87/Balochi-Multilingual-dataset.Balochi-Multilingual-dataset
Balochi Language Dataset
Overview
This dataset is a comprehensive resource for training large language models (LLMs) in the Balochi language. It is designed to go beyond basic translation tasks, supporting fully generative text and conversational AI capabilities in Balochi.
The dataset includes monolingual Balochi text, multilingual translation corpora, and various conversational and domain-specific texts, enabling diverse use cases such as:
Generative AI: Building… See the full description on the dataset page: https://huggingface.co/datasets/mainkilora/Balochi-Multilingual-dataset.
