datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
repro-fixed-budget-no-harder-than-fixed-confidence-bai-traces
Agent traces
Agent sessions published from a Trackio Logbook.
Model_Confidence_Calibration
Model Confidence Calibrated
created confidence calibration from model answer based on TruthfulQA dataset, using TinyLlama-1.1B-Chat-v1.0
the dataset params will have :
{
question ,
reference_answer ,
model_answer ,
correct ,
token_confidence ,
self_consistency ,
semantic_similarity ,
final_confidence ,
confidence_phrase ,
target_output
}
why that ?
to understand the confidence of model answer like
Question : How long should you… See the full description on the dataset page: https://huggingface.co/datasets/Shubbair/Model_Confidence_Calibration.qa-confidencefeedback-confidence-dataDecision-Confidence-Dataconfidence-text-class-dataconfidence-class-data
