PedroDKE/LibriS2S
LibriS2S This repo contains scripts and alignment data to create a dataset build further upon librivoxDeEn such that it contains (German audio, German transcription, English audio, English transcription) quadruplets and can be used for Speech-to-Speech translation research. Because of this, the alignments are released under the same Creative Commons Attribution-NonCommercial-ShareAlike 4.0 International License These alignments were collected by downloading the English… See the full description on the dataset page: https://huggingface.co/datasets/PedroDKE/LibriS2S.
update readme
update readme
update readme with a TOC and examples for usage
Merge branch 'main' of https://huggingface.co/datasets/PedroDKE/LibriS2S
update dataset_info, placement of ds file and test loader script
Upload dataset
testing of loading hs dataset
update dataset
update notebook
update dataset
update on dataset
add config
update dataset
use dl manager in dataset
update readme
moved file
update readme
update split
update dataset and dataset info
upload hf dataset and rename py dataset
Update README.md
Update README.md
Delete metadata.csv
Update README.md
Update README.md
Upload metadata.csv
Upload all_de_en_alligned_cleaned.csv
Update README.md
Update README.md
Upload 4 files
Delete alignments/clean_up_csv.py
Delete alignments/data_example.ipynb
Delete alignments/libris2s_dataset.py
uploaded 3 files
Upload all_de_en_alligned_cleaned.csv
Update README.md
update the tasks on the dataset card
Update README.md
update links in readme
Update README.md
Update README.md
German files
English files
Additional example files
Allignment files
example folder
Upload README.md
initial commit
