sac
Datasets
All datasets matching “sac”c-sac-corpora
C-SAC LibriTTS-R training subset
Deterministically selected and resampled speech from
mythicinfinity/libritts_r
for the C-SAC causal speech-codec program. The package retains source revision,
Parquet shard, row, utterance, transcript, and content hashes. LibriTTS-R is
distributed under CC BY 4.0; downstream users remain responsible for attribution.
Only prefixes with a hash-bound _COMPLETE.json sentinel are admissible.
tennis-sackmann-archive
Tennis datasets — archive of Jeff Sackmann's data
This dataset is an archival mirror of the public tennis datasets compiled by Jeff Sackmann.
It exists so the data remains available and citable. It contains only data and its documentation —
no models, analysis, or derived code.
A matching mirror lives on GitHub: https://github.com/Aneeshers/tennis-sackmann-archive
Contents
Folder
What it is
Files
Coverage
slam_pointbypoint/
Point-by-point logs for the… See the full description on the dataset page: https://huggingface.co/datasets/Aneeshers/tennis-sackmann-archive.sacrebleu_manualSACSoNSACRED-Bench
SACRED-Bench
This repository hosts SACRED-Bench (Speech-Audio Composition for RED-teaming), a benchmark designed to evaluate the robustness of Multimodal Large Language Models (LLMs) against complex audio-based attacks.
SACRED-Bench is introduced in the paper Speech-Audio Compositional Attacks on Multimodal LLMs and Their Mitigation with SALMONN-Guard.
Unlike existing perturbation-based methods, SACRED-Bench exploits speech-audio composition mechanisms to create challenging… See the full description on the dataset page: https://huggingface.co/datasets/tsinghua-ee/SACRED-Bench.berkeley_gnm_sac_son_raw
