MadvaAparna/KannadaVibhaktiSamples
Dataset Card for Dataset Name This is a dataset for the Named Entity Recognition (NER) task in Kannada language. Each instance represents one of the vibhakti cases. Dataset Details Dataset Description Kannada language supports eight (grammatical) cases [Refer https://kannadakalike.org/grammar/cases]. The cases are called vibhakti and corresponding suffixes, pratyaya. Noun words are inflected with the suffix corresponding to these cases to create… See the full description on the dataset page: https://huggingface.co/datasets/MadvaAparna/KannadaVibhaktiSamples.
Dataset Card for Dataset Name
This is a dataset for the Named Entity Recognition (NER) task in Kannada language. Each instance represents one of the vibhakti cases.
Dataset Details
Dataset Description
Kannada language supports eight (grammatical) cases [Refer https://kannadakalike.org/grammar/cases]. The cases are called vibhakti and corresponding suffixes, pratyaya. Noun words are inflected with the suffix corresponding to these cases to create another grammatically meaningful word.
In this dataset, 12 sentences are composed for each case, thus resulting in a total of 96 sentences. Each sentence contains at least one word representing any named entity which is inflected with the suffix for the corresponding case. We include examples for PER, LOC and ORG class of entities. For PER class, we also include entities which span across multiple words such as caMdragupta maurya. This vibhakti dataset can be used to analyse how the presence of case specific suffixes in a named entity affects the prediction by the model.
- Curated by: [More Information Needed]
- Funded by [optional]: [More Information Needed]
- Shared by [optional]: [More Information Needed]
- Language(s) (NLP): [More Information Needed]
- License: [More Information Needed]
Dataset Sources [optional]
<!-- Provide the basic links for the dataset. -->
- Repository: [More Information Needed]
- Paper [optional]: [More Information Needed]
- Demo [optional]: [More Information Needed]
Uses
Dataset can be used in case of the NER task.
Dataset Structure
<!-- This section provides a description of the dataset fields, and additional information about the dataset structure such as criteria used to create the splits, relationships between data points, etc. -->
[More Information Needed]
Source Data
Manually created
Who are the annotators?
<!-- This section describes the people or systems who created the annotations. -->
[More Information Needed]
