CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01sentence-transformers /example-documents Example Documents A small set of example documents across modalities (image, audio, video) for use in Sentence Transformers retrieval snippets and documentation. These are the kinds of files you pass to model.encode_document(...). They can safely be used as examples in your model cards if you don't want to host the example assets in your model repositories themselves. Contents File Modality doc1.jpg image (document page) doc2.jpg image (document page)… See the full description on the dataset page: https://huggingface.co/datasets/sentence-transformers/example-documents.audion<1K1 likes868 downloads1mo agoHugging Face02Derur /Derur-Translator-Examples Derur Translator Examples All complete examples (with all files created by the translator), including the originals.Все полные примеры (со всеми файлами, которые создаёт переводчик), включая оригиналы. More information about my translator: BoostyБольше информации о моём переводчике: Boosty audiotranslation0 likes780 downloads9mo agoHugging Face03hf-internal-testing /ashraq-esc50-1-dog-example Dataset Card for "ashraq-esc50-1-dog-example" More Information needed audion<1K0 likes720 downloads2y agoHugging Face04nccratliri /wing-flap-noise-audio-examplesaudion<1K0 likes679 downloads2y agoHugging Face05hf-internal-testing /dummy-flac-single-exampleaudion<1K0 likes362 downloads4y agoHugging Face06sarahwei /Taiwanese-Minnan-Example-Sentences Taiwanese Minnan Example Sentences The dataset consists of a collection of example sentences designed to aid in recognizing Taiwanese Minnan (Taiwanese Hokkien) for automatic speech recognition (ASR) tasks. This dataset is sourced from the Ministry of Education in Taiwan and aims to provide valuable linguistic resources for researchers and developers working on speech recognition systems. Dataset Features Source: Ministry of Education, Taiwan (Sutian Resource Center) Text:… See the full description on the dataset page: https://huggingface.co/datasets/sarahwei/Taiwanese-Minnan-Example-Sentences.audioautomatic-speech-recognition10K<n<100K12 likes284 downloads2y agoHugging Face07Narsil /candle-examplesaudion<1K0 likes266 downloads3y agoHugging Face08richardskimco /examplesaudion<1K0 likes114 downloads1y agoHugging Face09datasets-examples /doc-audio-6 [doc] audio dataset 6 This dataset contains 4 audio files in the /train directory, with a CSV metadata file providing another data column. audion<1K0 likes94 downloads2y agoHugging Face10ai-coustics /voice-focus-examples Voice Focus Examples Collection of examples for accurate foreground speaker transcription. Details Curated by: Joschka Wohlgemuth Funded by: ai-coustics GmbH Contact: Web: https://ai-coustics.com audion<1K2 likes78 downloads22d agoHugging Face11miracleyin /example_mmdata_mnbvc mnbvc mm dataset v2.1 MNBVC 多模态语料数据格式。原链接:https://huggingface.co/datasets/wanng/example_mmdata_mnbvc 参考实现:mm_template_mnbvc 的 mmdata_block.BLOCK_SCHEMA。schema 以那份代码为准,这个数据集是它的示例产物。 字段 字段名称 类型 字段说明 可选 实体ID string 数据的唯一标识符。用于在数据集中确定是哪一条数据。在单个数据集中确定一条数据的实体对象。 必选 md5 string 内容的 md5,用于去重与完整性校验 必选 块ID int32 一个实体对象内的标识符。用于确定一条数据内的一个部分数据。parquet 行的最小单元。 必选 块类型 string 用于保存块的类别。类别的含义为「模态」。取值见下 必选 扩展字段 string 用于保存块的元信息。为可以被成功 load 的 json 字符串。后期可继续扩展 必选… See the full description on the dataset page: https://huggingface.co/datasets/miracleyin/example_mmdata_mnbvc.audioimage-to-textn<1K2 likes72 downloads24d agoHugging Face12moonshine-ai /multilingual_examplesaudion<1K0 likes67 downloads1y agoHugging Face13bezzam /vocos-examplesaudion<1K0 likes67 downloads1y agoHugging Face14datasets-examples /doc-audio-1 [doc] audio dataset 1 This dataset contains 4 wav audio files at the root. audion<1K0 likes64 downloads2y agoHugging Face15foragi /Omni-DuplexEval-ExamplesOmni-DuplexEval-Examples is a curated subset of Omni-DuplexEval for qualitative visualization and paper demonstration. Each task contains 5 representative samples, covering all benchmark scenarios and task types. This subset is intended for illustrative purposes in the paper and supplementary materials. The annotation format and data structure are consistent with the full benchmark. Omni-DuplexEval Omni-DuplexEval is a benchmark for evaluating real-time duplex multimodal… See the full description on the dataset page: https://huggingface.co/datasets/foragi/Omni-DuplexEval-Examples.audion<1K0 likes62 downloads5mo agoHugging Face16eist-edinburgh /examplesaudion<1K1 likes47 downloads6mo agoHugging Face17naive-puzzle /commonvoice22-sidon-xcodec2-exampleaudio1K<n<10K0 likes31 downloads10mo agoHugging Face18Shrey160 /hinglish-audio-turn-end-example Hinglish Audio Turn End Example Dataset of hinglish audio snippets generated with SarvamAI bulbul:v3 TTS. Each example is a customer service utterance with metadata about the speaker gender, whether the utterance is complete or incomplete (trailoff / connective / filler ending), and the rendered audio file. Fields field type description script string hinglish transcript (Devanagari + Latin script mixed) audio audio mp3 speech clip generated by… See the full description on the dataset page: https://huggingface.co/datasets/Shrey160/hinglish-audio-turn-end-example.audion<1K0 likes31 downloads1mo agoHugging Face19kemuriririn /index-tts-2-examplesaudion<1K0 likes30 downloads1y agoHugging Face20minato-ryan /starrail-voice-snac-exampleaudio1K<n<10K0 likes29 downloads10mo agoHugging Face21datasets-examples /doc-audio-4 [doc] audio dataset 4 This dataset contains 2 wav audio files in the train/ subdirectory and 2 wav audio files in the test/ subdirectory. audion<1K0 likes24 downloads2y agoHugging Face22datasets-examples /doc-audio-2 [doc] audio dataset 2 This dataset contains 4 wav audio files in the audio/ subdirectory. audion<1K0 likes21 downloads2y agoHugging Face23donghao-zhou /OmniShow_example_dataset OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation Donghao Zhou1,*, Guisheng Liu2,*, Hao Yang2, Jiatong Li2,†, Jingyu Lin3, Xiaohu Huang4, Yichen Liu2, Xin Gao2, Cunjian Chen3, Shilei Wen2,§, Chi-Wing Fu1, Pheng-Ann Heng1,§ 1The Chinese University of Hong Kong, 2ByteDance, 3Monash University, 4The University of Hong Kong *Equal contribution, †Project lead, §Corresponding author 🌍 Useful Links Project Page:… See the full description on the dataset page: https://huggingface.co/datasets/donghao-zhou/OmniShow_example_dataset.audion<1K0 likes21 downloads5mo agoHugging Face24Anonymousv22222 /MuseBench-Exampleaudion<1K0 likes21 downloads5mo agoHugging Face25polinaeterna /audiofolder_exampleaudion<1K0 likes20 downloads4y agoHugging Face26Samhita /whisper-jax-examples Dataset Card for Dataset Name Dataset Summary This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields [More Information Needed] Data Splits [More Information Needed] Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/Samhita/whisper-jax-examples.audion<1K0 likes20 downloads3y agoHugging Face27nateraw /examplesaudion<1K0 likes19 downloads4y agoHugging Face28datasets-examples /doc-audio-3 [doc] audio dataset 3 This dataset contains 4 audio files with different formats (aiff, ogg, mp3 and flac) in the audio/ subdirectory. audion<1K0 likes19 downloads2y agoHugging Face29datasets-examples /doc-audio-9 [doc] audio dataset 9 This dataset contains four audio files, two in the /train directory (one in the cat/ subdirectory and one in the dog/ subdirectory), and two in the test/ directory (same distribution in subdirectories). audion<1K0 likes18 downloads2y agoHugging Face30minato-ryan /starrail-voice-xcodec2-exampleaudio1K<n<10K0 likes18 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.