FBK-MT/fama-data
Dataset Description, Collection, and Source The FAMA training data is the collection of English and Italian datasets for automatic speech recognition (ASR) and speech translation (ST) used to train the FAMA models family. The ASR section of FAMA is derived from the MOSEL data collection, including the automatic transcripts obtained with Whisper and available in the HuggingFace MOSEL Dataset. The ASR is further augmented with automatically transcribed speech from the… See the full description on the dataset page: https://huggingface.co/datasets/FBK-MT/fama-data.
Update README.md
Rename train_mls-it-en.tsv to train_mls_it-en.tsv
Add FAMA data
Add MT inference code to the main README
Correct YouTube-Commons README with the correct call to the scripts
Add segment-ytc.py script
Add speech_only.py script
Add YouTube-Commons ids and Silero logs
Create SplitAudioUsingSileroLog.pl
Add main FAMA data README
Add ratio filtering script for ST
Add YouTube-Commons README
initial commit
