nlpso/m0_fine_tuning_ocr_cmbert_io
m0_fine_tuning_ocr_cmbert_io Introduction This dataset was used to fine-tuned Jean-Baptiste/camembert-ner for flat NER task using Flat NER approach [M0]. It contains 19th-century Paris trade directories' entries. Dataset parameters Approach : M0 Dataset type : noisy (Pero OCR) Tokenizer : Jean-Baptiste/camembert-ner Tagging format : IO Counts : Train : 6084 Dev : 676 Test : 1685 Associated fine-tuned model : nlpso/m0_flat_ner_ocr_cmbert_io… See the full description on the dataset page: https://huggingface.co/datasets/nlpso/m0_fine_tuning_ocr_cmbert_io.
Upload README.md with huggingface_hub
Upload README.md with huggingface_hub
Upload data/test-00000-of-00001-c918c665dadb35fd.parquet with huggingface_hub
Upload data/dev-00000-of-00001-ef9d7b8215402f85.parquet with huggingface_hub
Upload data/train-00000-of-00001-5eab72fbf08eab89.parquet with huggingface_hub
initial commit
