lus
Datasets
All datasets matching “lus”ne-asr-dataset-lus-aug
NE ASR Augmented Dataset -- Mizo (lus)
Augmented automatic speech recognition dataset for Mizo (lus),
a Tibeto-Burman language spoken in Mizoram, India.
Source
Augmented from sulabhkatiyar/ne-asr-lus
(original transcribed speech data from the ARTPARK-IISc Vaani project).
Language Information
Property
Value
Language
Mizo
ISO 639-3
lus
Family
Tibeto-Burman
Region
Mizoram, India
Tonal
Yes
Tier
D (20.75h original data)… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-asr-dataset-lus-aug.ne-asr-dataset-lus
Mizo (lus) — ASR dataset
A small Mizo (lus) speech-to-text dataset for automatic speech recognition
(ASR) of a low-resource North-East India language. Each example pairs a short audio
clip with its Romanized (Latin-script) transcript.
Source
Derived from the ARTPARK-IISc Vaani project (https://vaani.iisc.ac.in/)
Splits
Split
Samples
train
9,850
validation
1,190
test
1,218
Data fields
Each example has:
audio — the… See the full description on the dataset page: https://huggingface.co/datasets/sulabhkatiyar/ne-asr-dataset-lus.dictionnaire_Lusignan
[!NOTE]
Dataset origin: http://www.dictionnaires-machtotz.org/index.php?option=com_content&view=article&id=89%3Adictionnaires-numerisesscannes&catid=44%3Aproduits-en-ligne&Itemid=74&lang=fr
Description
Le Nouveau Dictionnaire Illustré français-arménien de Guy de Lusignan, en deux volumes, est l’un des plus riches dictionnaires bilingues français-arménien occidental : Vol. 1, 1060 pages et Vol. 2, 817 pages. Il comporte près de 140 000 entrées en français.
anna-lush-krea2-datasetlusciousLuSNAR
