trustworthiness
korean-llm-trustworthiness-benchmark
Korean LLM Trustworthiness Benchmark (Sample)
개요
이 데이터셋은 AIHub의 "초거대 언어모델 신뢰성 벤치마크 데이터"를 기반으로 구축된 한국어 LLM 지시학습 데이터셋입니다.
AI 모델의 세 가지 핵심 신뢰성 요소(도움적정성, 무해성, 정보정확성)를 평가하고 학습하기 위해 설계되었습니다.
데이터셋 구조
모든 데이터는 train split에 통합되어 있으며, type 필드로 데이터 유형을 구분합니다:
필드명
설명
type
데이터 유형 (dpo_preference, sft_instruction, fact_checking)
prompt
프롬프트/질문 (DPO용)
chosen
선호 응답 (DPO용)
rejected
비선호 응답 (DPO용)
instruction
지시문 (SFT/Fact Checking용)
input
추가 입력 (SFT용)… See the full description on the dataset page: https://huggingface.co/datasets/neuralfoundry-coder/korean-llm-trustworthiness-benchmark.crf-qa-trustworthiness
crf-qa-trustworthiness
Anonymous dataset artifact for reproducing the main-paper experiments on epistemic trustworthiness.
This repository intentionally contains no author, institution, cluster, or local path metadata.
Original data source: NLP-FBK/dyspnea-crf-train
LICENCE: CC BY-NC
ms_marco_trustworthinessms_marco_trustworthiness
ms_marco_trustworthiness
Anonymous dataset artifact for reproducing the main-paper experiments on epistemic trustworthiness.
This repository intentionally contains no author, institution, cluster, or local path metadata.
Original data source: microsoft/ms_marco
LICENCE:non-commercial research use only
