CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01facebook /seamless-interaction Seamless Interaction Dataset A large-scale multimodal dataset of 4,000+ hours of human interactions for AI research 🖼️ Blog 🌐 Website 🎮 Demo 📦 GitHub 📄 Paper Human communication involves a complex interplay of verbal and nonverbal signals, essential for conveying meaning and achieving interpersonal goals. The Seamless Interaction Dataset is a large-scale collection of over 4,000 hours of face-to-face interaction footage from more than 4,000 participants in… See the full description on the dataset page: https://huggingface.co/datasets/facebook/seamless-interaction.audio198 likes107k downloads1y agoHugging Face02asahi417 /seamless-align-enA-hiAaudio100K<n<1M1 likes2.5k downloads2y agoHugging Face03rajjanardhan00 /Seamless_Dummy_Dataset_Fixed MMLU-Pro json This is a reupload of MMLU-Pro in json format. Please, refer to the original dataset for details. audioquestion-answeringn<1K0 likes1.4k downloads1y agoHugging Face04asahi417 /seamless-align-enA-jaAaudio100K<n<1M0 likes1.4k downloads2y agoHugging Face05asahi417 /seamless-align-enA-esAaudio1M<n<10M1 likes1.4k downloads2y agoHugging Face06asahi417 /seamless-align-deA-enAaudio100K<n<1M0 likes935 downloads2y agoHugging Face07kennethli319 /seamless-interaction-jefferson-annotations Seamless Interaction Jefferson-Style Annotations An automatic, turn-oriented annotation layer for the Meta Seamless Interaction Dataset. It compares the dataset's traditional transcript with an ASR-derived Jefferson-style condition and supplies speech-act, communicative-purpose, interactional-signal, alignment, and quality fields. This is a derived noncommercial research dataset. It does not redistribute the source audio. Every record retains the original interaction ID, split… See the full description on the dataset page: https://huggingface.co/datasets/kennethli319/seamless-interaction-jefferson-annotations.tabularautomatic-speech-recognition100K<n<1M0 likes869 downloads2mo agoHugging Face08asahi417 /seamless-align-enA-koAaudio100K<n<1M1 likes828 downloads2y agoHugging Face09ai4bharat /SeamlessAligngated BhasaAnuvaad: A Speech Translation Dataset for 13 Indian Languages Overview BhasaAnuvaad, is the largest Indic-language AST dataset spanning over 44,400 hours of speech and 17M text segments for 13 of 22 scheduled Indian languages and English. This repository consists of parallel data for Speech Translation from SeamlessAlign, a subset of BhasaAnuvaad. How to use The datasets library allows you to load and pre-process your dataset in pure Python… See the full description on the dataset page: https://huggingface.co/datasets/ai4bharat/SeamlessAlign.audio1M<n<10M6 likes731 downloads2y agoHugging Face10asahi417 /seamless-align-enA-frAaudio1M<n<10M0 likes703 downloads2y agoHugging Face11asahi417 /seamless-align-enA-viAaudio100K<n<1M0 likes556 downloads2y agoHugging Face12rajjanardhan00 /Seamless_Dummy_Dataset_Fixed_4license: cc-by-4.0 task_categories: object-detection video-classification tags: biology pretty_name: Seamless_Dummy audion<1K0 likes418 downloads1y agoHugging Face13asahi417 /seamless-align-enA-zhAaudio100K<n<1M3 likes338 downloads2y agoHugging Face14rajjanardhan00 /Seamless_Dummy_Dataset_Fixed_3 MMLU-Pro json This is a reupload of MMLU-Pro in json format. Please, refer to the original dataset for details. audioquestion-answeringn<1K0 likes266 downloads1y agoHugging Face15asahi417 /seamless-align-enA-jpnaudio100K<n<1M0 likes257 downloads2y agoHugging Face16SayantanJoker /processed_seamless_align_hindi_chunk_3audio10K<n<100K0 likes183 downloads1y agoHugging Face17SayantanJoker /processed_seamless_align_hindi_chunk_1audio10K<n<100K0 likes158 downloads1y agoHugging Face18SayantanJoker /processed_seamless_align_hindi_chunk_5audio10K<n<100K0 likes153 downloads1y agoHugging Face19SayantanJoker /processed_seamless_align_hindi_chunk_2audio10K<n<100K0 likes136 downloads1y agoHugging Face20SayantanJoker /processed_seamless_align_hindi_chunk_4audio10K<n<100K0 likes135 downloads1y agoHugging Face21SayantanJoker /processed_seamless_align_hindi_chunk_6audio10K<n<100K0 likes127 downloads1y agoHugging Face22asahi417 /seamless-align-enA-estaudio100K<n<1M0 likes114 downloads2y agoHugging Face23SayantanJoker /processed_seamless_align_hindi_chunk_10audio10K<n<100K0 likes99 downloads1y agoHugging Face24SayantanJoker /processed_seamless_align_hindi_chunk_19audio10K<n<100K0 likes70 downloads1y agoHugging Face25SayantanJoker /processed_seamless_align_hindi_chunk_20audio10K<n<100K0 likes69 downloads1y agoHugging Face26SayantanJoker /processed_seamless_align_hindiaudio1M<n<10M2 likes58 downloads1y agoHugging Face27AhmedBadawy11 /seamless_transcription_Cleaned_dataaudio1K<n<10K0 likes52 downloads2y agoHugging Face28zihan-audio /seamless-bg Seamless Background Robustness Pilot V3 expansion available: v3/README.md documents the expanded 1,985-event pool. Use v3/events_all.jsonl and v3/clips_all.jsonl for combined manifests. The original pilot statistics and files below remain unchanged. A compact, paired-audio candidate pool for incremental full-duplex interaction alignment and later background-speech augmentation. Derived from Meta's Seamless Interaction, by selecting events from the original train split only. This… See the full description on the dataset page: https://huggingface.co/datasets/zihan-audio/seamless-bg.audion<1K0 likes51 downloads16d agoHugging Face29frankie137 /aligned-seamless-interactionaudio10K<n<100K0 likes31 downloads5mo agoHugging Face30SayantanJoker /processed_seamless_align_hindi_new_chunk_49audio10K<n<100K0 likes30 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.