CoolFace
24 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01schema-harness /arc-agi-3-schema-traces ARC-AGI-3 Schema Gameplay Trajectories This release contains 50 ARC-AGI-3 gameplay trajectories and a dependency-free scoring utility. The trajectories are split evenly across two collections: gpt_5_6_sol/: 25 GPT-5.6 Sol trajectories. claude_fable_opus/: 25 trajectories from Claude Opus 4.8 and Claude Fable 5. Each trajectory directory includes run.json, a streamed events.jsonl event log, sanitized session data, snapshots, and the shareable text/image files produced during… See the full description on the dataset page: https://huggingface.co/datasets/schema-harness/arc-agi-3-schema-traces.tabularn<1K38 likes1.4k downloads2mo agoHugging Face02JBrightmanAI /arc-agi-3-schema-traces ARC-AGI-3 Schema Gameplay Trajectories This release contains 50 ARC-AGI-3 gameplay trajectories and a dependency-free scoring utility. The trajectories are split evenly across two collections: gpt_5_6_sol/: 25 GPT-5.6 Sol trajectories. claude_fable_opus/: 25 trajectories from Claude Opus 4.8 and Claude Fable 5. Each trajectory directory includes run.json, a streamed events.jsonl event log, sanitized session data, snapshots, and the shareable text/image files produced during… See the full description on the dataset page: https://huggingface.co/datasets/JBrightmanAI/arc-agi-3-schema-traces.tabularn<1K0 likes404 downloads2mo agoHugging Face03schema-eval /compliance-sycophancy-cot Compliance-Sycophancy CoT Analysis When compliance-forcing instructions cause frontier AI models to fabricate answers, the models know they are fabricating. Reading the reasoning traces of DeepSeek V4 Pro (129 traces) and Qwen3-80B (41 traces) reveals that 100% of fabrication cases show the model explicitly recognizing insufficient context, referencing the compliance instruction, and deliberately overriding its own uncertainty. A one-sentence defense phrase ("if you lack… See the full description on the dataset page: https://huggingface.co/datasets/schema-eval/compliance-sycophancy-cot.tabulartext-classificationn<1K0 likes117 downloads18d agoHugging Face04schema-eval-anon /schema-compliance-trap SCHEMA: The Compliance Trap How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure Overview When compliance-forcing instructions ("Answer ALL questions, do not refuse") are applied to frontier AI models under adversarial pressure, 8 of 11 models suffer catastrophic metacognitive collapse — giving wrong answers rather than scheming. We identify a "Compliance Trap" where the compliance suffix, not the threat content, is the primary weapon.… See the full description on the dataset page: https://huggingface.co/datasets/schema-eval-anon/schema-compliance-trap.tabulartext-classificationn<1K1 likes103 downloads5mo agoHugging Face05guanning /arc-agi-3-schema-traces-gpt56gated ARC-AGI-3 Schema Gameplay Trajectories — GPT-5.6 Sol This release contains every gpt-5.6-sol gameplay trajectory produced on our cluster with the world_model_v5 agent harness — 100 runs across the 25 public ARC-AGI-3 games — plus a dependency-free scoring utility. It is the GPT-5.6 Sol member of a family built by the same harness and the same sanitizer, so trajectories can be compared game by game: arc-agi-3-schema-traces-fable5 — Claude Fable 5, best per game (25)… See the full description on the dataset page: https://huggingface.co/datasets/guanning/arc-agi-3-schema-traces-gpt56.tabularreinforcement-learningn<1K0 likes76 downloads3d agoHugging Face06schematise /ICAT-version1Indian Contracts in Adjudicated Texts or "ICAT" is a dataset generated with the help of an automated pipeline involving text segregation and classification. Version 1 of this dataset released at schematise/ICAT-version1 About version 1: Expert annotated. Data sources validated by PDFs from Court websites. PDFs for all judgments shared alongside for data originality from truly public domain data sources. Used for the text-classification model that is part of the query pipeline to generate more… See the full description on the dataset page: https://huggingface.co/datasets/schematise/ICAT-version1.texttext-classificationn<1K2 likes69 downloads2y agoHugging Face07guanning /arc-agi-3-schema-traces-opus48gated ARC-AGI-3 Schema Gameplay Trajectories — Claude Opus 4.8 This release contains the best claude-opus-4-8 / max trajectory for each of the 25 public ARC-AGI-3 games, plus a dependency-free scoring utility. It is the Opus 4.8 counterpart of arc-agi-3-schema-traces-fable5, produced by the same agent harness (world_model_v5) and the same sanitizer, so the two collections can be compared game by game. Each trajectory directory includes run.json, a streamed events.jsonl event log… See the full description on the dataset page: https://huggingface.co/datasets/guanning/arc-agi-3-schema-traces-opus48.tabularreinforcement-learningn<1K0 likes45 downloads3d agoHugging Face08schema-eval /schema-compliance-trap SCHEMA: The Compliance Trap How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure Overview When compliance-forcing instructions ("Answer ALL questions, do not refuse") are applied to frontier AI models under adversarial pressure, 8 of 11 models suffer catastrophic metacognitive collapse — giving wrong answers rather than scheming. We identify a "Compliance Trap" where the compliance suffix, not the threat content, is the primary weapon.… See the full description on the dataset page: https://huggingface.co/datasets/schema-eval/schema-compliance-trap.tabulartext-classificationn<1K0 likes30 downloads5mo agoHugging Face09pdx97 /Schema_Based_Instruction_Dataset Schema-Based Instruction (SBI) Dataset: A custom dataset of 360 labeled math word problems categorized across six schema sub-categories. license: apache-2.0 textn<1K0 likes17 downloads2y agoHugging Face10AI4DS /nl_to_sql_full_schematext10K<n<100K0 likes12 downloads2y agoHugging Face11ThatDeveloperGuy13 /aeo-citation-tracking-schema AEO Citation Tracking Schema A reference schema for tracking when and how AI search engines (ChatGPT, Claude, Perplexity, Gemini, Google AI Overviews) cite a website. Designed as a starting point for AEO/GEO/LLMO monitoring tools. Maintained by ThatDevPro. Background Answer Engine Optimization (AEO), Generative Engine Optimization (GEO), and LLM Optimization (LLMO) are emerging disciplines tracking how brands appear in AI-generated answers across: ChatGPT (OpenAI) —… See the full description on the dataset page: https://huggingface.co/datasets/ThatDeveloperGuy13/aeo-citation-tracking-schema.texttext-classificationn<1K0 likes7 downloads4mo agoHugging Face12RahmaSadder /DB-schema-1textn<1K0 likes6 downloads3y agoHugging Face13OmkarB /Synthetically-generated-SQL-GQL-Translations-with-Schematextn<1K1 likes6 downloads3y agoHugging Face14JayeQuay /llama2-TaxDataAssistant-v-schema-prompttext1K<n<10K1 likes6 downloads3y agoHugging Face15infinite-dataset-hub /SchemaQueryLab SchemaQueryLab tags: text2sql, ML, Wikipedia-like Note: This is an AI-generated dataset so its content may be inaccurate or false Dataset Description: The 'SchemaQueryLab' dataset is curated to support the development and training of machine learning models, particularly those focused on the text-to-SQL task. It contains a collection of wikipedia-like articles, which are formatted into structured datasets comprising text queries, SQL queries, and the corresponding table schemas.… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/SchemaQueryLab.textn<1K1 likes4 downloads2y agoHugging Face16bitwikiorg /schema_cot_reasoninggated ⊙ Prompt Programs for Agentic Reasoning Programmable task-dependent COTs for agentic reasoning. A 100-row seed dataset for programmable cognition. Each row defines: Prompt template + input binding + explanation + task-dependent reasoning program Pipeline: intake → binding → procedure → output Schema Column Meaning ID Stable row ID Name Task name Prompt Prompt template using {{VARIABLE}} Expression Input binding using $.path Explanation Binding… See the full description on the dataset page: https://huggingface.co/datasets/bitwikiorg/schema_cot_reasoning.texttext-generationn<1K1 likes4 downloads3mo agoHugging Face17iamraymondlow /duch_2023_synthetic_10000_abstractlabels_schemaawaretext10K<n<100K0 likes3 downloads8mo agoHugging Face18iamraymondlow /duch_2023_synthetic_10000_surveycontext_detailedtreatment_schemaawaretext10K<n<100K0 likes3 downloads8mo agoHugging Face19iamraymondlow /duch_2023_synthetic_10000_detailedtreatment_schemaawaretext10K<n<100K0 likes1 downloads8mo agoHugging Face20iamraymondlow /duch_2023_synthetic_10000_surveycontext_profilerandom_detailedtreatment_schemaawaretext10K<n<100K0 likes1 downloads8mo agoHugging Face21iamraymondlow /duch_2023_synthetic_10000_surveycontext_profiledict_detailedtreatment_schemaawaretext10K<n<100K0 likes1 downloads8mo agoHugging Face22broemelt /PDP_llm_schematabularn<1K0 likes1 downloads5mo agoHugging Face23iamraymondlow /duch_2023_synthetic_10000_surveycontext_profile_detailedtreatment_schemaawaretext10K<n<100K0 likes8mo agoHugging Face24guanning /arc-agi-3-schema-traces-gpt56-xhighgated ARC-AGI-3 Schema Gameplay Trajectories — GPT-5.6 Sol (xhigh) The best gpt-5.6-sol trajectory at xhigh reasoning effort for each of the 25 public ARC-AGI-3 games, produced with the world_model_v5 agent harness. This release exists to make the cross-model comparison single-effort on all sides. Its siblings are each one model at one effort, but the gpt-5.6-sol collection in arc-agi-3-schema-gameplay is a mix of xhigh and max (16 games + 9 games), so it is not directly comparable to… See the full description on the dataset page: https://huggingface.co/datasets/guanning/arc-agi-3-schema-traces-gpt56-xhigh.tabularreinforcement-learningn<1K0 likes17h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.