empathyai/books-ner-dataset
Books Named Entity Recognition (NER) Dataset A lightweight Named‑Entity‑Recognition (NER) corpus built from titles and author names contained in Project Gutenberg’s public catalogues. It is intended for training or benchmarking entity extractors such as Gliner on bibliographic metadata. 1 Provenance This dataset provenance originates from Project Gutenberg's public catalogue. 2 Quick facts Records (total) 434 925… See the full description on the dataset page: https://huggingface.co/datasets/empathyai/books-ner-dataset.
224
No commit history came back for main. The revision may not exist, or the source declined the request.
