CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01NCSOFT /Designed-Vocalizations-Dataset Designed Vocalizations Dataset Paper · Demo & audio samples The Designed Vocalizations Dataset supports voice conversion for designed vocalizations — monster growls, robotic voices, and other sound-designed timbres — an area left underexplored by benchmarks that focus on natural human speech. It curates diverse raw vocal sources (speech and animal / non-linguistic sounds) and applies professional vocal-effects processing to produce corresponding effect-modified variants. A… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/Designed-Vocalizations-Dataset.audioaudio-to-audio100K<n<1M4 likes759 downloads2mo agoHugging Face02NCSOFT /K-MMBench K-MMBench We introduce K-MMBench, a Korean adaptation of the MMBench [1] designed for evaluating vision-language models. By translating the dev subset of MMBench into Korean and carefully reviewing its naturalness through human inspection, we developed a novel robust evaluation benchmark specifically for Korean language. K-MMBench consists of questions across 20 evaluation dimensions, such as identity reasoning, image emotion, and attribute recognition, allowing a thorough… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/K-MMBench.image1K<n<10K15 likes718 downloads1y agoHugging Face03NCSOFT /K-MMStar K-MMStar We introduce K-MMStar, a Korean adaptation of the MMStar [1] designed for evaluating vision-language models. By translating the val subset of MMStar into Korean and carefully reviewing its naturalness through human inspection, we developed a novel robust evaluation benchmark specifically for Korean language. (We observe that there are unanswerable cases (e.g., multiple images required to answer the question but only has a single image, vague questions or options) in the… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/K-MMStar.image1K<n<10K12 likes441 downloads1y agoHugging Face04NCSOFT /K-SEED K-SEED We introduce K-SEED, a Korean adaptation of the SEED-Bench [1] designed for evaluating vision-language models. By translating the first 20 percent of the test subset of SEED-Bench into Korean, and carefully reviewing its naturalness through human inspection, we developed a novel robust evaluation benchmark specifically for Korean language. K-SEED consists of questions across 12 evaluation dimensions, such as scene understanding, instance identity, and instance attribute… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/K-SEED.image1K<n<10K23 likes369 downloads1y agoHugging Face05NCSOFT /K-DTCBench K-DTCBench We introduce K-DTCBench, a newly developed Korean benchmark featuring both computer-generated and handwritten documents, tables, and charts. It consists of 80 questions for each image type and two questions per image, summing up to 240 questions in total. This benchmark is designed to evaluate whether vision-language models can process images in different formats and be applicable for diverse domains. All images are generated with made-up values and statements for… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/K-DTCBench.imagen<1K16 likes274 downloads1y agoHugging Face06NCSOFT /offsetbias Dataset Card for OffsetBias Dataset Description: 💻 Repository: https://github.com/ncsoft/offsetbias 📜 Paper: OffsetBias: Leveraging Debiased Data for Tuning Evaluators Dataset Summary OffsetBias is a pairwise preference dataset intended to reduce common biases inherent in judge models (language models specialized in evaluation). The dataset is introduced in paper OffsetBias: Leveraging Debiased Data for Tuning Evaluators. OffsetBias contains 8,504 samples… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/offsetbias.texttext-classification1K<n<10K27 likes136 downloads2y agoHugging Face07NCSOFT /K-LLaVA-W K-LLaVA-W We introduce K-LLaVA-W, a Korean adaptation of the LLaVA-Bench-in-the-wild [1] designed for evaluating vision-language models. By translating the LLaVA-Bench-in-the-wild into Korean and carefully reviewing its naturalness through human inspection, we developed a novel robust evaluation benchmark specifically for Korean language. (Since our goal was to build a benchmark exclusively focused in Korean, we change the English texts in images into Korean for localization.)… See the full description on the dataset page: https://huggingface.co/datasets/NCSOFT/K-LLaVA-W.imagen<1K16 likes81 downloads1y agoHugging Face08GenRM /offsetbias-NCSOFT Dataset Card for OffsetBias Dataset Description: 💻 Repository: https://github.com/ncsoft/offsetbias 📜 Paper: OffsetBias: Leveraging Debiased Data for Tuning Evaluators Dataset Summary OffsetBias is a pairwise preference dataset intended to reduce common biases inherent in judge models (language models specialized in evaluation). The dataset is introduced in paper OffsetBias: Leveraging Debiased Data for Tuning Evaluators. OffsetBias contains 8,504 samples… See the full description on the dataset page: https://huggingface.co/datasets/GenRM/offsetbias-NCSOFT.texttext-classification1K<n<10K0 likes15 downloads1y agoHugging Face09open-llm-leaderboard /NCSOFT__Llama-VARCO-8B-Instruct-detailsgated Dataset Card for Evaluation run of NCSOFT/Llama-VARCO-8B-Instruct Dataset automatically created during the evaluation run of model NCSOFT/Llama-VARCO-8B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NCSOFT__Llama-VARCO-8B-Instruct-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face10math-extraction-comp /NCSOFT__Llama-VARCO-8B-Instructtabular1K<n<10K0 likes7 downloads2y agoHugging Face11PJMixers /NCSOFT_offsetbias-PreferenceShareGPTtext1K<n<10K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.