CoolFace
Datasetpublic

RolandM/imprecision-bench

imprecision-bench A multimodal benchmark for evaluating whether LLMs calibrate linguistic precision to pragmatic context, paired with 475 human productions and a peer-reviewed RSA baseline (r² ≈ 0.97). This dataset accompanies the paper: Modeling (Im)precision in Context Roland Mühlenbernd, Stephanie Solt Linguistics Vanguard, 2022 [Paper] · [Source Data] · [Companion Repo] Notebook notebook.ipynb — guided walkthrough: data loading, sample evaluation (1 row… See the full description on the dataset page: https://huggingface.co/datasets/RolandM/imprecision-bench.

sourceHugging Facecc-by-4.0updated 1mo agoView on Hugging Face
0likes98downloads
settings

This repository belongs to RolandM on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameimprecision-bench
visibilitypublic
licencecc-by-4.0
gatedno
ownerRolandM
Account settings
RolandM/imprecision-bench · CoolFace