local-dataset
movie_stills_captioned_dataset_local
Dataset Card for "movie_stills_captioned_dataset_local"
More Information needed
bio-vla-datasets-local-sync-20260618
argo11/bio-vla-datasets-local-sync-20260618
概要
Bio-VLA 実験環境で利用した dataset 群の local sync snapshot です。主に LIBERO benchmark の robot manipulation demonstration HDF5 と、関連するデータ同期成果物を含みます。
確認した内容
ファイル数: 8,292
合計サイズ: 約 499 GB
主な構成:
libero/libero_10/*.hdf5
libero/libero_90/*.hdf5
task 名は KITCHEN_SCENE..., LIVING_ROOM_SCENE..., STUDY_SCENE... などの language-conditioned manipulation task です。
外部根拠
LIBERO は lifelong robot learning のための robot… See the full description on the dataset page: https://huggingface.co/datasets/argo11/bio-vla-datasets-local-sync-20260618.thai-local-language-translation-dataset
Thai Local Language Translation Dataset
Thai Local Language Translation Dataset is a translation dataset for translate Thai Local Language to Thai Central Language. We create the dataset from Thai Dialect Corpus (Thai dialects ASR corpus). We select train set only from Thai Dialect Corpus.
The dataset support Khummuang, Korat, and Pattani.
Reference
Suwanbandit, A., Naowarat, B., Sangpetch, O., Chuangsuwanich, E. (2023) Thai Dialect Corpus and Transfer-based Curriculum… See the full description on the dataset page: https://huggingface.co/datasets/pythainlp/thai-local-language-translation-dataset.Bangali_local_dialect_ASR_HF_Dataset
BanglaMix — Code-Switching ASR in Bangladeshi Regional Dialects
BanglaMix is a speech dataset for Automatic Speech Recognition (ASR) on dialectal Bangladeshi Bengali mixed with English (code-switching). It covers 15 regional dialects and the natural Bengali–English code-switching common in informal Bangladeshi speech — a setting not covered by existing Bengali corpora, which address either dialects or code-switching, never both.
Clips
41,499 transcribed audio clips… See the full description on the dataset page: https://huggingface.co/datasets/niloycste68/Bangali_local_dialect_ASR_HF_Dataset.faceberg-dataset-localdexter-local-pm-3b-dataset
