CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01CG-Bench /CG-Benchgated CG-Bench Project Website: https://cg-bench.github.io/leaderboard/GitHub Repository: https://github.com/CG-Bench/CG-Bench (includes running code) Summary We introduce CG-Bench, a groundbreaking benchmark for clue-grounded question answering in long videos, addressing the limitations of existing benchmarks that focus primarily on short videos and rely on multiple-choice questions (MCQs). These limitations allow models to answer by elimination rather than genuine… See the full description on the dataset page: https://huggingface.co/datasets/CG-Bench/CG-Bench.tabularvisual-question-answering10K<n<100K9 likes10k downloads1y agoHugging Face02shuzhig /cgbench CG-Bench (mini) — clue-grounded long-video QA A local mirror of the CG-Bench mini split: 3,000 multiple-choice questions over 1,118 long videos (mean length ~28 min), each question annotated with the clue intervals (second-level time spans) in the video that actually justify the answer. Contents Path Size Description cgbench_mini.json 2.3 MB 3,000 QA items (see schema below) durations.json 38 KB {video_uid: duration_in_seconds} for 1,219 videos… See the full description on the dataset page: https://huggingface.co/datasets/shuzhig/cgbench.tabularvideo-text-to-text1K<n<10K0 likes2k downloads1mo agoHugging Face03cgato /SlimOrcaDedupCleaned What is this dataset? Half of the Slim Orca Deduped dataset, but further cleaned by removing instances of soft prompting. I removed a ton prompt prefixes which did not add any information or were redundant. Ex. "Question:", "Q:", "Write the Answer:", "Read this:", "Instructions:" I also removed a ton of prompt suffixes which were simply there to lead the model to answer as expected Ex. "The answer is...", "Answer:", "A:", "Summary:", "Output:", "Highlight:" Why? I… See the full description on the dataset page: https://huggingface.co/datasets/cgato/SlimOrcaDedupCleaned.text100K<n<1M27 likes814 downloads3y agoHugging Face04CG-Bench /CG-AV-Countinggated CG-AV-Counting Updates [2025/07/22] Since errors in a few clue annotations when converting frame indexes to timestamps, there were errors in the previous benchmark leaderboard, we have reevaluated all models and have updated the new leaderboard. Summary Despite progress in video understanding, current MLLMs struggle with counting tasks. Existing benchmarks are limited by short videos, close-set queries, lack of clue annotations, and weak… See the full description on the dataset page: https://huggingface.co/datasets/CG-Bench/CG-AV-Counting.textvisual-question-answering1K<n<10K5 likes250 downloads1y agoHugging Face05DukeNLP /tailor-cgo Dataset Card for Tailor-CGO This dataset contains evaluations of language-model-generated responses regarding vaccine concerns, where each response is tailored to establish common ground through an identified "Common-Ground Opinion". Dataset Details Dataset Description The dataset contains both human- and LLM-annotated preferences/scores for how "well tailored" each written response is. Annotations are structured as a (1) relative preference between two… See the full description on the dataset page: https://huggingface.co/datasets/DukeNLP/tailor-cgo.texttext-generation10K<n<100K2 likes85 downloads2y agoHugging Face06ApyHTML19 /Cgi_Impots_Marocaine_2026text1K<n<10K1 likes76 downloads16d agoHugging Face07CGICAI /cherokee-english-translation Cherokee–English Parallel Corpus (Archivist Project) A curated Cherokee (ᏣᎳᎩ / Tsalagi) ↔ English parallel corpus for machine translation, assembled from public sources, deduplicated, benchmark-decontaminated, and conflict-cleaned. Built to train and evaluate English→Cherokee translation models for one of the most endangered languages in North America. Files File Rows Purpose train_en2chr_v2.jsonl 138,307 Flagship training set. English→Cherokee SFT… See the full description on the dataset page: https://huggingface.co/datasets/CGICAI/cherokee-english-translation.tabulartranslation100K<n<1M0 likes56 downloads2mo agoHugging Face08cga-bench-neurips26 /cga-bench CGA-Bench Hugging Face Collection This dataset repo is a collection index for the nine reviewer-facing CGA-Bench dataset descriptors used in the NeurIPS 2026 E&D submission. Included configs overview: collection-level summary row spanning the full benchmark release main_corpus: 19,062-episode primary evaluation corpus source_grounded: source-grounded SGSC subset graph_anchored: graph-anchored SGSC subset profile_expanded: profile-expanded SGSC subset auto_expanded: 76… See the full description on the dataset page: https://huggingface.co/datasets/cga-bench-neurips26/cga-bench.tabulartext-generationn<1K0 likes47 downloads5mo agoHugging Face09Adam1010 /cgrt-consensus-5model CGRT Consensus 5-Model Dataset Multi-model consensus dataset for studying model agreement and disagreement patterns on mathematical reasoning tasks. Dataset Description 61,678 math problems evaluated by 5 frontier LLMs with full reasoning traces and extracted answers. Models Used Model Provider Version Claude Anthropic claude-3-5-sonnet-20241022 Codex/GPT-4 OpenAI gpt-4o Gemini Google gemini-1.5-flash DeepSeek DeepSeek deepseek-chat Qwen… See the full description on the dataset page: https://huggingface.co/datasets/Adam1010/cgrt-consensus-5model.tabularquestion-answering10K<n<100K0 likes46 downloads9mo agoHugging Face10louisbrulenaudet /cgi Code Général des Impôts, non-instruct (11-12-2023) This project focuses on fine-tuning pre-trained language models to create efficient and accurate models for tax practice. Fine-tuning is the process of adapting a pre-trained model to perform specific tasks or cater to particular domains. It involves adjusting the model's parameters through a further round of training on task-specific or domain-specific data. While conventional fine-tuning strategies involve supervised learning… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/cgi.texttext-generation1K<n<10K1 likes41 downloads3y agoHugging Face11cg48 /ndn-cottext1K<n<10K0 likes40 downloads1y agoHugging Face12Keith13th /cgc-registrytextn<1K0 likes39 downloads26d agoHugging Face13cgulse /alpaca-cleaned-trAlpaca Cleaned Dataset. Machine Translated facebook/nllb-200-3.3B Languages Turkish text10K<n<100K2 likes33 downloads3y agoHugging Face14CGIAR /gardian-ai-ready-docsgated ⚠️ Heads up: Updated Dataset Available This dataset has been updated with a newer version published on 27 Feb 2025. The latest version includes more updated and refined set of documents. We recommend using the latest version, available at https://huggingface.co/datasets/CGIAR/gardian-cigi-ai-documents. This version remains accessible for reference and reproducibility purposes. A Curated Research Corpus for Agricultural Advisory AI Applications This dataset… See the full description on the dataset page: https://huggingface.co/datasets/CGIAR/gardian-ai-ready-docs.textsummarization10K<n<100K2 likes26 downloads9mo agoHugging Face15cg48 /ndn-adv-cottextn<1K0 likes26 downloads1y agoHugging Face16cg48 /ndn-synth-adv-qual-qacurrently broken in a lot of places textn<1K0 likes24 downloads1y agoHugging Face17jameschennerd /CG-bench CG-Bench: A Call Graph Construction Benchmark for Language Models 📄 Paper: CG-Bench: Can Language Models Assist Call Graph Construction in the Real World? - Published at LMPL@SPLASH2025 CG-Bench is a comprehensive benchmark dataset designed to evaluate the capabilities of Large Language Models (LLMs) in assisting with call graph construction in real-world C/C++ codebases. The benchmark focuses specifically on challenging indirect function calls through function pointers, which… See the full description on the dataset page: https://huggingface.co/datasets/jameschennerd/CG-bench.documentn<1K0 likes17 downloads1y agoHugging Face18cgato /TheSmarts8/1 - Added Hermes 3 Dataset A mixture of synthetic data pulled from all over HuggingFace and then aggressively deduplicated. Nearly 20gb of synth data crunched down to just under 2Gb. Also cleaned up any system prompts which would not be expected to alter behavior. Mostly done as an excercise in deduplication and ngram analysis. Used to finetune https://huggingface.co/cgato/Nemo12b-TheSyntheticOne If you use this dataset you need to mask out the samples labeled false, not doing so will… See the full description on the dataset page: https://huggingface.co/datasets/cgato/TheSmarts.texttext-generation100K<n<1M2 likes14 downloads1y agoHugging Face19cg48 /ndn-basic-qatextn<1K0 likes13 downloads1y agoHugging Face20cg48 /testtextn<1K0 likes12 downloads1y agoHugging Face21CG80499 /Inverse-scaling-testtabularmultiple-choice1K<n<10K0 likes11 downloads4y agoHugging Face22Lux0926 /MetaMath-Mistral-7B-CGPO-10ktabular10K<n<100K0 likes11 downloads11mo agoHugging Face23cg48 /ndn-adv-qual-2textn<1K0 likes9 downloads1y agoHugging Face24PJMixers /AP-News-2024-CGPT-Summarize-ShareGPTThe AP News dataset, run through ChatGPT (gpt-3.5-turbo) to get summaries. All use the same system prompt; "You summarize text. Ensure that your summaries effectively capture key points, while being concise." Currently not all of the articles from the dataset are summarized, since I keep hitting "You've reached our limit of messages per hour. Please try again later." texttext-generationn<1K1 likes8 downloads3y agoHugging Face25open-llm-leaderboard /cgato__TheSalt-L3-8b-v0.3.2-detailsgated Dataset Card for Evaluation run of cgato/TheSalt-L3-8b-v0.3.2 Dataset automatically created during the evaluation run of model cgato/TheSalt-L3-8b-v0.3.2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/cgato__TheSalt-L3-8b-v0.3.2-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face26cgsharefive /traintextn<1K0 likes8 downloads2y agoHugging Face27Lux0926 /DeepSeekMath-Base-7B-SFT-CGPO-10ktabular10K<n<100K0 likes8 downloads11mo agoHugging Face28Lux0926 /Deepseek-Coder-7B-Instruct-v1.5-CGPO-10ktabular10K<n<100K0 likes8 downloads11mo agoHugging Face29CGTec3 /wiki-kieWiki頁面截圖VQA資料集 隨機Wiki頁面截圖的VQA資料集 格式如下 [ { "from": "human", "value": "<image>請提取圖片中的文字資訊,並以JSON格式返回\"條目名稱\"、\"宣傳標語\"和\"內容摘要\",遵照格式 {\"條目名稱\": \"\", \"宣傳標語\": \"\", \"內容摘要\": \"\"}" }, { "from": "gpt", "value": "{\"條目名稱\": \"長吻眶鋸雀鯛\", \"宣傳標語\": \"MoWiki維基編輯定期聚每月第三個星期六於台中舉辦,歡迎報名參加和關注我們。\", \"內容摘要\": \"長吻眶鋸雀鯛,又稱鈍頭高身雀鯛,俗名為厚殼仔,為輻鰭魚綱鱸形目雀鯛科的其中一種。\"}" } ] 為避免噪聲,標記僅涵蓋條目名稱、宣傳標語與內容摘要;建議不要使用宣傳標語,過濾後再進行訓練。 Json檔案無遵照常見標記結構,請根據檔案路徑進行匹配… See the full description on the dataset page: https://huggingface.co/datasets/CGTec3/wiki-kie.image1K<n<10K0 likes7 downloads11mo agoHugging Face30Lux0926 /MetaMath-Llama-8B-CGPO-10ktabular10K<n<100K0 likes6 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.