datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
chaos-mnli-ambiguityChaos NLI MNLI portion with gini coefficient pre-computed (from 0 to 1)
High gini means unambiguous inference.
@inproceedings{xzhou2022distnli,
Author = {Xiang Zhou and Yixin Nie and Mohit Bansal},
Booktitle = {Findings of the Association for Computational Linguistics: ACL 2022},
Publisher = {Association for Computational Linguistics},
Title = {Distributed NLI: Learning to Predict Human Opinion Distributions for Language Reasoning},
Year = {2022}
}
ambiguity-casebook
Dual-Use Ambiguity Casebook
A 35-row, single-annotator research corpus for studying context-sensitive
adjudication in AI-mediated biology. Records follow one exact 21-field schema
and cover six descriptive categories.
This is a dataset release, not a model leaderboard, compliance tool, or
laboratory guide. Raw model responses, model-comparison results, live-provider
evaluation code, and adversarial failure-mode material are intentionally
excluded.
Dataset structure… See the full description on the dataset page: https://huggingface.co/datasets/jang1563/ambiguity-casebook.http-status-ambiguity-guide
HTTP Status Code Guide for Ambiguous API Scenarios
This dataset provides guidance on selecting appropriate HTTP status codes in common, yet ambiguous, API development situations. It helps developers make consistent and semantically correct choices, improving API predictability and client integration.
25 rows · category: reference · licence: CC0-1.0 (public domain)
Usage
import http_status_guide # Assuming the module is saved as http_status_guide.py
scenarios =… See the full description on the dataset page: https://huggingface.co/datasets/SharkSkin/http-status-ambiguity-guide.dissent-synthetic-clause-ambiguity
Dissent — synthetic financial clauses with seeded formalization ambiguity
Code · Interactive Space
Every clause in this dataset is synthetic. No real contract, client document or
third-party text was used, quoted or paraphrased. The language follows standard market
forms of drafting so that it is representative of the constructions that cause real
formalization disputes.
What this is for
Autoformalization research is crowded at the producing end and empty at the… See the full description on the dataset page: https://huggingface.co/datasets/NagaYu/dissent-synthetic-clause-ambiguity.AmbiguityDataset
Ambiguity Resolution Dataset
Overview
This dataset contains 25,656 samples for training and evaluating ambiguity resolution capabilities in robot navigation and interaction systems. It covers common object reference ambiguities in indoor scenes.
Why This Dataset
Real human instructions are often vague, incomplete, or inconsistent with the environment, but existing VLN datasets assume perfect clarity.
This dataset introduces realistic ambiguity so that models… See the full description on the dataset page: https://huggingface.co/datasets/qiqiquq/AmbiguityDataset.han-domestic-task-ambiguity-benchmark-v1
Domestic Task Ambiguity Benchmark
A benchmark dataset designed to analyze
how ambiguity in human instructions
affects task understanding in humanoid robots.
Methods
Instructions are annotated based on
explicitness and reference clarity.
Use Cases
Ambiguity handling research
Instruction disambiguation
Limitations
No multi-step task chains included.
Part of
Humanoid Network (HAN)
License
MIT
Ambiguity_Detector
