CoolFace
Datasetpublic

willem640/MAAT_documentary_with_DDbDP_and_HGV_metadata

This is dataset containing the documentary papyri from MAAT, with metadata from the HGV and DDbDP. The scripts in preprocessing/maat were used to generate these files. The processing is not perfect, but this should have metadata for most papyri where it was available. The fields corpus_id, file_id, block_index, id, title, material, language, training_text and test_cases were derived from MAAT: Fitzgerald, W., & Barney, J. (2024). The Machine-Actionable Ancient Text (MAAT) Corpus (1.0.0-beta)… See the full description on the dataset page: https://huggingface.co/datasets/willem640/MAAT_documentary_with_DDbDP_and_HGV_metadata.

sourceHugging Facecc-by-4.0updated 4mo agoView on Hugging Face
0likes16downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
willem640/MAAT_documentary_with_DDbDP_and_HGV_metadata · CoolFace