datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AgentChat-Test
Test Set Description
This directory contains the test set used for tool-use evaluation. The JSON files under Test-JSON/ are organized by task type:
SingleTaskProcessing/tool-select_test.json: single-tool selection tasks.
ParallelProcessing/parallel-call_test.json: parallel tool-call tasks.
ProactiveSeeking/searchTools_test_predictions_kept.json: proactive tool-search tasks.
TaskDecomposition/muti-tool-select_test.json: multi-tool task decomposition tasks.… See the full description on the dataset page: https://huggingface.co/datasets/leungtianle/AgentChat-Test.MutiEmo-Test
MultiEmo-Test
MultiEmo-Test is an English evaluation set for instruction-following multi-emotion text-to-speech synthesis. It accompanies HybridEmo, a system for modeling sequential emotion trajectories and simultaneous emotion blending within an utterance.
The dataset is intended for evaluation only. It contains synthesis text, natural-language emotion instructions, emotion annotations, and prompt audio for speaker-timbre conditioning. It does not contain target synthesized… See the full description on the dataset page: https://huggingface.co/datasets/ICTNLP/MutiEmo-Test.test321
test321
This is a merged speech dataset containing 118 audio segments from 2 source datasets.
Dataset Information
Total Segments: 118
Speakers: 4
Languages: tr
Emotions: happy, angry, sad, neutral
Original Datasets: 2
Dataset Structure
Each example contains:
audio: Audio file (WAV format, 16kHz sampling rate)
text: Transcription of the audio
speaker_id: Unique speaker identifier (made unique across all merged datasets)
emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/test321.hafidh-test-fixtureslaser-vibrations-test
Laser Vibrations
Dataset of laser speckle vibration recordings used to locate objects hidden inside a cardboard box.
A 10×10 grid of lasers shines on the side of a box containing an object; as loudspeakers excite the box,
the speckle patterns shift in proportion to the local surface vibration. Per-sample metadata is in
data/metadata.jsonl; full signal data and media files live in per-sample subdirectories.
Dataset Viewer Columns
Column
Type
Description… See the full description on the dataset page: https://huggingface.co/datasets/eturok-weizmann/laser-vibrations-test.Nsynth_Test_Split_Tango_Formatwhisper-finetune-audio_test2testdataset
NeMo Tarred Dataset
Generated from Test3.
Train rows: 40338 · Test rows: 422
Shards: 5 · Codec: flac · Sample rate: 16000 Hz mono
Primary text: text · target_lang: ta-IN
is_tarred: true
tarred_audio_filepaths: .../audio__OP_0..4_CL_.tar
manifest_filepath: .../train_manifest.json
whisper-finetune-audio_test3testtr43
testtr43
This is a merged speech dataset containing 2655 audio segments from 3 source datasets.
Dataset Information
Total Segments: 2655
Speakers: 13
Languages: tr
Emotions: angry, happy, neutral
Original Datasets: 3
Dataset Structure
Each example contains:
audio: Audio file (WAV format, original sampling rate preserved)
text: Transcription of the audio
speaker_id: Unique speaker identifier (made unique across all merged datasets)
emotion: Detected emotion… See the full description on the dataset page: https://huggingface.co/datasets/Codyfederer/testtr43.
