didi0di/KoWoW
Dataset Card for KoWoW Dataset Summary WoW(Wiard of Wikipedia)를 한국어로 변역한 데이터입니다. Dataset Description WoW(Wiard of Wikipedia)라는 지식 기반 대화 데이터를 한국어로 변역한 데이터입니다.한 대화에 여러 개의 dialog가 묶음으로 구성되어 있으며, 전체 대화는 22,311건, 전체 dialog는 201,999개 입니다.본 데이터셋은 Knowledge와 Utterance가 모두 한국어인 ko 버전만 가져온 데이터입니다. Language(s) (NLP): ko License: mit Dataset Sources Repository: https://github.com/AIRC-KETI/kowow/tree/master Uses… See the full description on the dataset page: https://huggingface.co/datasets/didi0di/KoWoW.
Dataset Card for KoWoW
<!-- Provide a quick summary of the dataset. -->
Dataset Summary
WoW(Wiard of Wikipedia)를 한국어로 변역한 데이터입니다.
Dataset Description
WoW(Wiard of Wikipedia)라는 지식 기반 대화 데이터를 한국어로 변역한 데이터입니다. 한 대화에 여러 개의 dialog가 묶음으로 구성되어 있으며, 전체 대화는 22,311건, 전체 dialog는 201,999개 입니다. 본 데이터셋은 Knowledge와 Utterance가 모두 한국어인 ko 버전만 가져온 데이터입니다.
- Language(s) (NLP): ko
- License: mit
Dataset Sources
<!-- Provide the basic links for the dataset. -->
- Repository: https://github.com/AIRC-KETI/kowow/tree/master
Uses
<!-- Address questions around how the dataset is intended to be used. -->
Source Data
<!-- This section describes the source data (e.g. news text and headlines, social media posts, translated sentences, ...). --> WoW(Wiard of Wikipedia)
Who are the source data producers?
<!-- This section describes the people or systems who originally created the data. It should also include self-reported demographic or identity information for the source data creators if this information is available. -->
- AIRC-KETI
