datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
h4p7t3x2-jn6b9_tran
affectexpect/h4p7t3x2-jn6b9_tran
This dataset contains transcribed audio files organized in folders for scalability.
Dataset Structure
The dataset is organized with:
Audio files: Stored in audio_XXXXX/ folders (5000 files per folder)
Metadata: Stored in data_XXXXX/ folders as parquet files
This organization follows Hugging Face best practices for datasets with millions of files.
Statistics
Total files: 926
Total batches: 2427
Audio folders: 3
Files per… See the full description on the dataset page: https://huggingface.co/datasets/affectexpect/h4p7t3x2-jn6b9_tran.t9p3c8m1-axr4e6_tran
affectexpect/t9p3c8m1-axr4e6_tran
This dataset contains transcribed audio files organized in folders for scalability.
Dataset Structure
The dataset is organized with:
Audio files: Stored in audio_XXXXX/ folders (5000 files per folder)
Metadata: Stored in data_XXXXX/ folders as parquet files
This organization follows Hugging Face best practices for datasets with millions of files.
Statistics
Total files: 8,901
Total batches: 5183
Audio folders: 6
Files per… See the full description on the dataset page: https://huggingface.co/datasets/affectexpect/t9p3c8m1-axr4e6_tran.n4x7d2q9-hf1m8t3_tran
affectexpect/n4x7d2q9-hf1m8t3_tran
This dataset contains transcribed audio files organized in folders for scalability.
Dataset Structure
The dataset is organized with:
Audio files: Stored in audio_XXXXX/ folders (5000 files per folder)
Metadata: Stored in data_XXXXX/ folders as parquet files
This organization follows Hugging Face best practices for datasets with millions of files.
Statistics
Total files: 922
Total batches: 11336
Audio folders: 10… See the full description on the dataset page: https://huggingface.co/datasets/affectexpect/n4x7d2q9-hf1m8t3_tran.
