l3cube-pune/marathi-pos-tagger
L3Cube-MahaPOS: Marathi Part-of-Speech Tagging Dataset Dataset Description L3Cube-MahaPOS is one of the first large-scale, manually annotated Part-of-Speech (POS) tagging datasets for Marathi — an Indo-Aryan language spoken by over 83 million people. The dataset comprises 32,354 sentences sourced from Marathi news text and annotated with a 16-tag scheme aligned with the Universal Dependencies (UD) v2 framework. This dataset is part of the L3Cube-MahaNLP family of… See the full description on the dataset page: https://huggingface.co/datasets/l3cube-pune/marathi-pos-tagger.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face