datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
named-entity-recognition
Sinhala Named Entity Recognition
Sinhala Named Entity Recognition is a token-level named entity recognition dataset for Sinhala.
This repository is a re-upload of the original Sinhala NER dataset introduced by Manamini et al. (2016) in "Ananya - a Named-Entity-Recognition (NER) System for Sinhala Language" with proper train/ test splits. The dataset was subsequently included as the Named Entity Recognition (NER) task in the SINHALA-GLUE benchmark introduced in "Sinhala… See the full description on the dataset page: https://huggingface.co/datasets/sinhala-nlp/named-entity-recognition.named_entity_recognitionlaw_entity_recognition
Dataset Card for Dataset Name
The dataset transforms complex legal passages into structured outputs, detailing entities, their interrelationships, and claims, providing a foundation for a legal knowledge graph to facilitate advanced analysis and applications.
Dataset Details
Dataset Description
The dataset in question is a specialized collection designed for legal text analysis, where each input is a passage of legal text—ranging from case law to statutory… See the full description on the dataset page: https://huggingface.co/datasets/rubenamtz0/law_entity_recognition.
