CoolFace
Datasetpublic

juletxara/lindsea-blimp

LINDSEA BLIMP Dataset Description LINDSEA BLIMP is a dataset of Indonesian linguistic minimal pairs for evaluating language models' syntactic knowledge. The dataset is based on the BHASA project's Indonesian syntax data. Dataset Structure The dataset contains minimal pairs of grammatical and ungrammatical sentences in Indonesian, organized into separate subsets testing various linguistic phenomena: argument_structure (160 pairs): Tests word order… See the full description on the dataset page: https://huggingface.co/datasets/juletxara/lindsea-blimp.

sourceHugging Facecc-by-4.0updated 1y agoView on Hugging Face
0likes68downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
juletxara/lindsea-blimp · CoolFace