CoolFace
13 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01automated-alignment-science /cot-controllability-steering-vectors CoT-controllability steering vectors — artifacts Artifacts for the project "A 2,880-number steering vector gives a reasoning model the chain-of-thought control that fine-tuning does" on gpt-oss-20b. Code + the master notebook + generate_figures.py that load these artifacts: https://github.com/redwoodresearch/automated-research-projects (folder cot-controllability-steering-vectors). Contents steering_vectors/ — the headline frozen-weights steering vector… See the full description on the dataset page: https://huggingface.co/datasets/automated-alignment-science/cot-controllability-steering-vectors.0 likes73 downloads3mo agoHugging Face02ejcgan /cot-controllability-steering-vectors CoT-controllability steering vectors — artifacts Artifacts for the project "Activation steering can increase chain-of-thought controllability" on gpt-oss-20b: a single frozen-weights steering vector (2,880 numbers added to one layer's residual stream) matches what a LoRA fine-tune does to the model's CoT controllability on held-out instructions, and works by raising the late attention heads' attention onto the in-context instruction. Code + the master notebook +… See the full description on the dataset page: https://huggingface.co/datasets/ejcgan/cot-controllability-steering-vectors.0 likes61 downloads2mo agoHugging Face03brendanlong /cot-controllability-gpt-oss-20b CoT-controllability elicitation on gpt-oss-20b — traces & soft prompts Raw artifacts for the experiment cot-controllability-experiment (full writeup and code there): can we prompts to control a model's chain of though by being louder and more detailed, by learning soft prompts, and by projecting those soft prompts back to hard prompts? Headline: Our hard prompts (even loud "dakka" rewrites) give ~0 control; a soft prompt works across 7 behaviours / 6 categories at 56–82%… See the full description on the dataset page: https://huggingface.co/datasets/brendanlong/cot-controllability-gpt-oss-20b.1K<n<10K0 likes44 downloads2mo agoHugging Face04Reih02 /cot_control_controllability_2000tabular1K<n<10K0 likes25 downloads6mo agoHugging Face05retkowski /length-controllability-evaluation Evaluation: Zero-Shot Strategies for Length-Controllable Summarization This repository contains summaries generated using various approaches and parameters, as part of a comprehensive study on length-controllable summarization with zero-shot methods. The summaries were created to evaluate LLMs' length control capabilities across multiple measures and to test practical methods for improving controllability. We refer to the paper (arXiv|aclanthology), presented as Findings paper at… See the full description on the dataset page: https://huggingface.co/datasets/retkowski/length-controllability-evaluation.summarization1 likes17 downloads1y agoHugging Face06Reih02 /cot_control_deepseek_v3_controllability_2000tabular1K<n<10K0 likes16 downloads6mo agoHugging Face07Reih02 /cot_control_kimi_k2_controllability_2000tabular1K<n<10K0 likes11 downloads6mo agoHugging Face08ClarusC64 /fusion-hybrid-controllability-loss-horizon-and-control-routing-v0.1What this dataset tests Forecasts when a hybrid reactor will lose controllable power. Maps stability margin collapse into a time horizon. Routes optimal intervention. Required outputs time_to_control_loss_min recommended_action Use case Predictive stabilization for subcritical or hybrid reactors. Prevents loss of controllability before shutdown conditions. tabulartabular-regressionn<1K0 likes7 downloads8mo agoHugging Face09Reih02 /cot_control_qwen3_32b_controllability_2000tabular1K<n<10K0 likes7 downloads6mo agoHugging Face10Reih02 /cot_control_kimi_k25_controllability_2000tabular1K<n<10K0 likes6 downloads6mo agoHugging Face11Snooow1029 /itts-controllabilitygated CommonVoice ITTS Controllability Validation Set Ground-truth-labelled English speech clips drawn from fixie-ai/common_voice_17_0 (Common Voice 17.0, CC0), used to validate audio-LM judges for instructable-TTS (ITTS) controllability scoring on age, gender, and native accent. Each clip carries a human-provided attribute label from Common Voice, so a judge's perceived-attribute accuracy can be measured against real ground truth (blind perceived-accuracy protocol).… See the full description on the dataset page: https://huggingface.co/datasets/Snooow1029/itts-controllability.audioaudio-classificationn<1K0 likes5 downloads2mo agoHugging Face12Reih02 /cot_control_qwen35_35b_controllability_20000 likes4 downloads6mo agoHugging Face13Reih02 /cot_control_gpt_oss_20b_controllability_2000tabular1K<n<10K0 likes4 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.