espnet/Bagpiper_SFT_Data
Bagpiper SFT Data Release status: the validated Parquet release is being uploaded. The homepage and metadata may appear before every large shard is committed. Bagpiper SFT Data is the supervised fine-tuning corpus for Bagpiper, an open-ended audio language model that understands and generates speech, music, environmental sound, and their mixtures through rich textual captions and planning. The public release has exactly two configurations: Configuration Direction… See the full description on the dataset page: https://huggingface.co/datasets/espnet/Bagpiper_SFT_Data.
16.6k
1{2 "audio_bytes": 1663142302282,3 "dataset": "bagpiper",4 "errors": [],5 "parquet_bytes": 1132362679743,6 "partitions": [7 {8 "audio_bytes": 411563826060,9 "parquet_bytes": 363645830327,10 "partition": "generation",11 "rows": 1474011,12 "shards": 1944,13 "source_subsets": {14 "part2_gen_v1_imaginary": 371827,15 "part2_gen_v1_realistic": 463850,16 "part3_gen_v1_imaginary": 72508,17 "part3_gen_v1_realistic": 95134,18 "part4_gen_v1_imaginary": 202108,19 "part4_gen_v1_realistic": 26858420 }21 },22 {23 "audio_bytes": 1251578476222,24 "parquet_bytes": 768716849416,25 "partition": "understanding",26 "rows": 1193006,27 "shards": 4860,28 "source_subsets": {29 "airbench_train_v1": 357896,30 "asr_v2_inverse_200k": 200000,31 "audiobench_train_v1": 300886,32 "mmau_train_v1": 33422433 }34 }35 ],36 "rows": 2667017,37 "schema_version": "bagpiper-self-contained-parquet-v1",38 "status": "passed",39 "unique_row_ids": 266701740}41 