cnc
Datasets
All datasets matching “cnc”cn_college_listen_mcq_test@article{wang2024audiobench,
title={AudioBench: A Universal Benchmark for Audio Large Language Models},
author={Wang, Bin and Zou, Xunlong and Lin, Geyu and Sun, Shuo and Liu, Zhuohan and Zhang, Wenyu and Liu, Zhengyuan and Aw, AiTi and Chen, Nancy F},
journal={NAACL},
year={2025}
}
CNCv2
[!NOTE]
This repository integrates the "V2" release of the Causal News Corpus (CNC) — published as RECESS — into hf
datasets. Please find the original dataset here. This is the
actively-maintained release the maintainers recommend using ("For 2023 Shared Task, please use V2"), with
far richer span annotations (2257 causal relations) than CNC, the original 2022 release (183
causal relations) — kept as its own separate dataset for comparison rather than silently overwritten.
Please see the… See the full description on the dataset page: https://huggingface.co/datasets/thagen/CNCv2.CNCF-Stuff
Model Card for Model ID
Model Details
Model Description
Developed by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Model type: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Finetuned from model [optional]: [More Information Needed]
Model Sources [optional]
Repository: [More Information Needed]
Paper… See the full description on the dataset page: https://huggingface.co/datasets/cwielenb/CNCF-Stuff.MuCE-Pref
Multi-task Creativity Evaluation Dataset - Preference subset (MuCE-Pref)
This dataset is introduced in the Creative Preference Optimization and contains human responses and ratings for multiple creativity assessments.
It is derived from the MuCE dataset and is in the TRL Preference dataset format.
Dataset Sources
See Creative Preference Optimization for a list of sources.
Citation
@misc{ismayilzada2025creativepreferenceoptimization… See the full description on the dataset page: https://huggingface.co/datasets/CNCL-Penn-State/MuCE-Pref.cncf-raw-data-for-llm-training
CNCF Raw Data for LLM Training
Description
This dataset, named cncf-raw-data-for-llm-training, consists of markdown (MD) and PDF content extracted from various project repositories within the CNCF (Cloud Native Computing Foundation) landscape. The data was collected by fetching MD and PDF files from different CNCF project repositories and converting them into JSON format. This dataset is intended as raw data for training large language models (LLMs).
The dataset includes… See the full description on the dataset page: https://huggingface.co/datasets/Kubermatic/cncf-raw-data-for-llm-training.MuCE
Multi-task Creativity Evaluation Dataset (MuCE)
This dataset is introduced in the Creative Preference Optimization and contains human responses and ratings for multiple creativity assessments.
Dataset Sources
See Creative Preference Optimization for a list of sources.
Citation
@misc{ismayilzada2025creativepreferenceoptimization,
title={Creative Preference Optimization},
author={Mete Ismayilzada and Antonio Laverghetta Jr. and Simone A. Luchini and… See the full description on the dataset page: https://huggingface.co/datasets/CNCL-Penn-State/MuCE.
