decoding
Datasets
All datasets matching “decoding”DecodingTrust-Agent-Platform
DecodingTrust-Agent Platform
A Controllable and Interactive Red-Teaming Platform for AI Agents.
This is the per-task dataset for the DecodingTrust-Agent Platform (DTAP),
spanning 14 real-world domains and 50+ simulation environments that replicate widely-used
systems such as Google Workspace, PayPal, Slack, Salesforce, Snowflake, and Databricks. Each task
ships the configuration the evaluator needs to spin up the sandbox, run an agent, and verify the
outcome — config.yaml (task… See the full description on the dataset page: https://huggingface.co/datasets/AI-Secure/DecodingTrust-Agent-Platform.brain-decodingDecodingTrust
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
Overview
This repo contains the source code of DecodingTrust. This research endeavor is designed to help researchers better understand the capabilities, limitations, and potential risks associated with deploying these state-of-the-art Large Language Models (LLMs). See our paper for details.
DecodingTrust: A Comprehensive Assessment of Trustworthiness in GPT Models
Boxin Wang, Weixin Chen, Hengzhi… See the full description on the dataset page: https://huggingface.co/datasets/AI-Secure/DecodingTrust.DLM-Decoding-Analysis
DLM-Decoding-Analysis
Diffusion Language Model Knows the Answer Before It Decodes
Pengxiang Li*, Yefan Zhou*, Dilxat Muhtar, Lu Yin, Shilin Yan, Li Shen, Yi Liang, Soroush Vosoughi, Shiwei Liu
The Fourteenth International Conference on Learning Representations (ICLR 2026)
TL;DR: Diffusion language models often commit to the correct answer
well before they finish decoding. This dataset releases the per-question,
step-by-step decoding trajectories of LLaDA-8B-Instruct on… See the full description on the dataset page: https://huggingface.co/datasets/YefanZhou98/DLM-Decoding-Analysis.speculative_decoding_benchmarksbrain-decoding-nsd
