taiwanese
Taiwanese-Minnan-Sutiau
Taiwanese-Minnan-Sutiau Dataset
The dataset consists of a curated collection of words that resemble tokens in Taiwanese Minnan (Taiwanese Hokkien), aimed at enhancing the recognition and processing of the language for various applications. Sourced from the Ministry of Education in Taiwan, this dataset serves as a valuable linguistic resource for researchers and developers engaged in language processing and recognition tasks.
Dataset Features
Source: Ministry of Education, Taiwan… See the full description on the dataset page: https://huggingface.co/datasets/sarahwei/Taiwanese-Minnan-Sutiau.Taiwanese-Minnan-Example-Sentences
Taiwanese Minnan Example Sentences
The dataset consists of a collection of example sentences designed to aid in recognizing Taiwanese Minnan (Taiwanese Hokkien) for automatic speech recognition (ASR) tasks. This dataset is sourced from the Ministry of Education in Taiwan and aims to provide valuable linguistic resources for researchers and developers working on speech recognition systems.
Dataset Features
Source: Ministry of Education, Taiwan (Sutian Resource Center)
Text:… See the full description on the dataset page: https://huggingface.co/datasets/sarahwei/Taiwanese-Minnan-Example-Sentences.Taiwanese-Chinese_characters-POJ-Collectionannotations_creators:
expert-generated
language:
zh
en
language_creators:
expert-generated
license:
mit
multi-linguality:
monolingual
pretty_name: '
Taiwanese text dataset: a Chinese characters and POJ collection'
size_categories:
1M<n<10M
source_datasets:
original
tags: []
task_categories:
text-classification
feature-extraction
task_ids:
multi-label-classification
multi-class-classification
taiwanese_english_translationThis new dataset is designed to solve this great NLP task and is crafted with a lot of care.Taiwanese_ASRcv-taiwanese
