Flaglab/esnlir-human-validated
ESNLIR — human-validated subset 972 sentence pairs from ESNLIR whose label was confirmed by human annotators — the validated subset used in An Analysis of the Performance of Large Language Models in Spanish NLI Datasets with Causal Relationships (IBERAMIA 2026, to appear). Part of the ESNLIR-LLM collection, used in Pacolas/NLI-via-LLM. A random sample of 2,136 instances from the full corpus was labeled by 27 Spanish-speaking university students, one label per pair. Only pairs… See the full description on the dataset page: https://huggingface.co/datasets/Flaglab/esnlir-human-validated.
Soften citation wording
Foreground the IBERAMIA 2026 paper these splits were built for
Cite the IBERAMIA 2026 paper; drop stale confusion-matrix reference
Set license to CC BY 4.0
Fix collection link
Add dataset card
Add labeled_final_dataset.jsonl
initial commit
