Sh1man/golos_opus
Dataset Description GOLOS is a Russian dataset for speech research. This is the OPUS version of the GOLOS dataset. Usage from datasets import load_dataset, Audio dataset = load_dataset("Sh1man/golos_opus", "crowd", split="train") print(dataset[0]['opus']) Dataset Statistics Dataset structure Domain Train files Train hours Test files Test hours Crowd 979 796 1 095 9 994 11.2 Farfield 124 003 132.4 1 916 1.4 Total 1 103… See the full description on the dataset page: https://huggingface.co/datasets/Sh1man/golos_opus.
Update README.md
Add files using upload-large-folder tool
Delete transcript
Delete golos_opus.py
Delete n_shards.json
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update golos_opus.py
Upload golos_opus.py
Delete golos_opus.py
Update golos_opus.py
Update golos_opus.py
Update README.md
Delete dataset_infos
Update README.md
Upload folder using huggingface_hub
Delete configs
Delete dataset_info.json
Upload dataset_info.json with huggingface_hub
Upload configs/train_farfield_config.json with huggingface_hub
Upload dataset
Upload configs/train_crowd_config.json with huggingface_hub
Upload dataset
Upload dataset
Upload dataset
Upload dataset
Upload dataset
Upload configs/test_farfield_config.json with huggingface_hub
Upload dataset
Upload configs/test_crowd_config.json with huggingface_hub
Upload dataset
Update README.md
Update README.md
initial commit
