CoolFace
Datasetpublic

CAMeL-Lab/BAREC-Corpus-v1.0

BAREC Corpus v1.0 Dataset Summary BAREC (the Balanced Arabic Readability Evaluation Corpus) is a large-scale dataset for fine-grained Arabic readability assessment. The dataset includes over 1M words, annotated at the sentence level across 19 readability levels, with additional mappings to coarser 7, 5, and 3 level schemes. Supported Tasks The dataset supports multi-class readability classification in the following formats: 19 levels (default) 7… See the full description on the dataset page: https://huggingface.co/datasets/CAMeL-Lab/BAREC-Corpus-v1.0.

sourceHugging Facecc-by-sa-4.0updated 1y agoView on Hugging Face
2likes184downloads

CAMeL-Lab/BAREC-Corpus-v1.0 · main · files are served by the source, never re-hosted here