CoolFace
Datasetpublic

l3cube-pune/marathi-pos-tagger

L3Cube-MahaPOS: Marathi Part-of-Speech Tagging Dataset Dataset Description L3Cube-MahaPOS is one of the first large-scale, manually annotated Part-of-Speech (POS) tagging datasets for Marathi — an Indo-Aryan language spoken by over 83 million people. The dataset comprises 32,354 sentences sourced from Marathi news text and annotated with a 16-tag scheme aligned with the Universal Dependencies (UD) v2 framework. This dataset is part of the L3Cube-MahaNLP family of… See the full description on the dataset page: https://huggingface.co/datasets/l3cube-pune/marathi-pos-tagger.

sourceHugging Facecc-by-4.0updated 2mo agoView on Hugging Face
0likes55downloads
4 commits on main
ddc879c2mo ago

Update README.md

l3cube-pune
0c134a63mo ago

Update README.md

l3cube-pune
16e91d93mo ago

Upload 3 files

l3cube-pune
8142e813mo ago

initial commit

l3cube-pune