kmi-linguistics/ilist
Dataset Card for ilist Dataset Summary This dataset is introduced in a task which aimed at identifying 5 closely-related languages of Indo-Aryan language family: Hindi (also known as Khari Boli), Braj Bhasha, Awadhi, Bhojpuri and Magahi. These languages form part of a continuum starting from Western Uttar Pradesh (Hindi and Braj Bhasha) to Eastern Uttar Pradesh (Awadhi and Bhojpuri) and the neighbouring Eastern state of Bihar (Bhojpuri and Magahi). For this task… See the full description on the dataset page: https://huggingface.co/datasets/kmi-linguistics/ilist.
Convert dataset to Parquet (#4)
Delete legacy JSON metadata (#3)
Replace YAML keys from int to str (#2)
Reorder split names (#1)
add dataset_info in dataset metadata
remove dummmy data
fix task_ids
Fix missing tags in dataset cards (#4896)
Fix titles in dataset cards (#4824)
Add `language_bcp47` tag (#4753)
Align more metadata with other repo types (models,spaces) (#4607)
Update files from the datasets library (from 1.18.0)
Update files from the datasets library (from 1.8.0)
Update files from the datasets library (from 1.7.0)
Update files from the datasets library (from 1.6.1)
Update files from the datasets library (from 1.6.0)
Update files from the datasets library (from 1.3.0)
Update files from the datasets library (from 1.2.0)
