datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
seamless-interaction
Seamless Interaction Dataset
A large-scale multimodal dataset of 4,000+ hours of human interactions for AI research
🖼️ Blog
🌐 Website
🎮 Demo
📦 GitHub
📄 Paper
Human communication involves a complex interplay of verbal and nonverbal signals, essential for conveying meaning and achieving interpersonal goals.
The Seamless Interaction Dataset is a large-scale collection of over 4,000 hours of face-to-face interaction footage from more than 4,000 participants in… See the full description on the dataset page: https://huggingface.co/datasets/facebook/seamless-interaction.seamless-align-enA-hiASeamless_Dummy_Dataset_Fixed
MMLU-Pro json
This is a reupload of MMLU-Pro in json format. Please, refer to the original dataset for details.
seamless-align-enA-jaAseamless-align-enA-esAseamless-align-deA-enAseamless-interaction-jefferson-annotations
Seamless Interaction Jefferson-Style Annotations
An automatic, turn-oriented annotation layer for the
Meta Seamless Interaction Dataset.
It compares the dataset's traditional transcript with an ASR-derived
Jefferson-style condition and supplies speech-act, communicative-purpose,
interactional-signal, alignment, and quality fields.
This is a derived noncommercial research dataset. It does not redistribute
the source audio. Every record retains the original interaction ID, split… See the full description on the dataset page: https://huggingface.co/datasets/kennethli319/seamless-interaction-jefferson-annotations.seamless-align-enA-koASeamlessAlign
BhasaAnuvaad: A Speech Translation Dataset for 13 Indian Languages
Overview
BhasaAnuvaad, is the largest Indic-language AST dataset spanning over 44,400 hours of speech and 17M text segments for 13 of 22 scheduled Indian languages and English.
This repository consists of parallel data for Speech Translation from SeamlessAlign, a subset of BhasaAnuvaad.
How to use
The datasets library allows you to load and pre-process your dataset in pure Python… See the full description on the dataset page: https://huggingface.co/datasets/ai4bharat/SeamlessAlign.seamless-align-enA-frAseamless-align-enA-viASeamless_Dummy_Dataset_Fixed_4license: cc-by-4.0
task_categories:
object-detection
video-classification
tags:
biology
pretty_name: Seamless_Dummy
seamless-align-enA-zhASeamless_Dummy_Dataset_Fixed_3
MMLU-Pro json
This is a reupload of MMLU-Pro in json format. Please, refer to the original dataset for details.
seamless-align-enA-jpnprocessed_seamless_align_hindi_chunk_3processed_seamless_align_hindi_chunk_1processed_seamless_align_hindi_chunk_5processed_seamless_align_hindi_chunk_2processed_seamless_align_hindi_chunk_4processed_seamless_align_hindi_chunk_6seamless-align-enA-estprocessed_seamless_align_hindi_chunk_10processed_seamless_align_hindi_chunk_19processed_seamless_align_hindi_chunk_20processed_seamless_align_hindiseamless_transcription_Cleaned_dataseamless-bg
Seamless Background Robustness Pilot
V3 expansion available: v3/README.md documents the expanded 1,985-event pool. Use v3/events_all.jsonl and v3/clips_all.jsonl for combined manifests. The original pilot statistics and files below remain unchanged.
A compact, paired-audio candidate pool for incremental full-duplex interaction alignment and later background-speech augmentation. Derived from Meta's Seamless Interaction, by selecting events from the original train split only. This… See the full description on the dataset page: https://huggingface.co/datasets/zihan-audio/seamless-bg.aligned-seamless-interactionprocessed_seamless_align_hindi_new_chunk_49
