CoolFace
Datasetpublic

espnet/Bagpiper_SFT_Data

Bagpiper SFT Data Release status: the validated Parquet release is being uploaded. The homepage and metadata may appear before every large shard is committed. Bagpiper SFT Data is the supervised fine-tuning corpus for Bagpiper, an open-ended audio language model that understands and generates speech, music, environmental sound, and their mixtures through rich textual captions and planning. The public release has exactly two configurations: Configuration Direction… See the full description on the dataset page: https://huggingface.co/datasets/espnet/Bagpiper_SFT_Data.

sourceHugging Faceupdated 2mo agoView on Hugging Face
1likes6.6kdownloads
VALIDATION_REPORT.json41 linesDownload Raw Back to root
1{2  "audio_bytes": 1663142302282,3  "dataset": "bagpiper",4  "errors": [],5  "parquet_bytes": 1132362679743,6  "partitions": [7    {8      "audio_bytes": 411563826060,9      "parquet_bytes": 363645830327,10      "partition": "generation",11      "rows": 1474011,12      "shards": 1944,13      "source_subsets": {14        "part2_gen_v1_imaginary": 371827,15        "part2_gen_v1_realistic": 463850,16        "part3_gen_v1_imaginary": 72508,17        "part3_gen_v1_realistic": 95134,18        "part4_gen_v1_imaginary": 202108,19        "part4_gen_v1_realistic": 26858420      }21    },22    {23      "audio_bytes": 1251578476222,24      "parquet_bytes": 768716849416,25      "partition": "understanding",26      "rows": 1193006,27      "shards": 4860,28      "source_subsets": {29        "airbench_train_v1": 357896,30        "asr_v2_inverse_200k": 200000,31        "audiobench_train_v1": 300886,32        "mmau_train_v1": 33422433      }34    }35  ],36  "rows": 2667017,37  "schema_version": "bagpiper-self-contained-parquet-v1",38  "status": "passed",39  "unique_row_ids": 266701740}41