empathyai/books-ner-dataset
Books Named Entity Recognition (NER) Dataset A lightweight Named‑Entity‑Recognition (NER) corpus built from titles and author names contained in Project Gutenberg’s public catalogues. It is intended for training or benchmarking entity extractors such as Gliner on bibliographic metadata. 1 Provenance This dataset provenance originates from Project Gutenberg's public catalogue. 2 Quick facts Records (total) 434 925… See the full description on the dataset page: https://huggingface.co/datasets/empathyai/books-ner-dataset.
211
No card is published for this repository, or it could not be fetched from Hugging Face right now.
