CoolFace
8 results

trustworthiness

neuralfoundry-coder /korean-llm-trustworthiness-benchmark Korean LLM Trustworthiness Benchmark (Sample) 개요 이 데이터셋은 AIHub의 "초거대 언어모델 신뢰성 벤치마크 데이터"를 기반으로 구축된 한국어 LLM 지시학습 데이터셋입니다. AI 모델의 세 가지 핵심 신뢰성 요소(도움적정성, 무해성, 정보정확성)를 평가하고 학습하기 위해 설계되었습니다. 데이터셋 구조 모든 데이터는 train split에 통합되어 있으며, type 필드로 데이터 유형을 구분합니다: 필드명 설명 type 데이터 유형 (dpo_preference, sft_instruction, fact_checking) prompt 프롬프트/질문 (DPO용) chosen 선호 응답 (DPO용) rejected 비선호 응답 (DPO용) instruction 지시문 (SFT/Fact Checking용) input 추가 입력 (SFT용)… See the full description on the dataset page: https://huggingface.co/datasets/neuralfoundry-coder/korean-llm-trustworthiness-benchmark.texttext-generation1K<n<10K0 likes20 downloads8mo agoHugging Faceanonymousubmission /crf-qa-trustworthiness crf-qa-trustworthiness Anonymous dataset artifact for reproducing the main-paper experiments on epistemic trustworthiness. This repository intentionally contains no author, institution, cluster, or local path metadata. Original data source: NLP-FBK/dyspnea-crf-train LICENCE: CC BY-NC text1K<n<10K0 likes11 downloads4mo agoHugging Faceferrazzipietro /ms_marco_trustworthinesstext1K<n<10K0 likes8 downloads5mo agoHugging Faceanonymousubmission /ms_marco_trustworthiness ms_marco_trustworthiness Anonymous dataset artifact for reproducing the main-paper experiments on epistemic trustworthiness. This repository intentionally contains no author, institution, cluster, or local path metadata. Original data source: microsoft/ms_marco LICENCE:non-commercial research use only text1K<n<10K0 likes8 downloads4mo agoHugging Face