CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01espnet /Bagpiper_SFT_Data Bagpiper SFT Data Release status: the validated Parquet release is being uploaded. The homepage and metadata may appear before every large shard is committed. Bagpiper SFT Data is the supervised fine-tuning corpus for Bagpiper, an open-ended audio language model that understands and generates speech, music, environmental sound, and their mixtures through rich textual captions and planning. The public release has exactly two configurations: Configuration Direction… See the full description on the dataset page: https://huggingface.co/datasets/espnet/Bagpiper_SFT_Data.audioaudio-classification1M<n<10M1 likes6.6k downloads2mo agoHugging Face02espnet /floras FLORAS FLORAS is a 50-language benchmark For LOng-form Recognition And Summarization of spoken language. The goal of FLORAS is to create a more realistic benchmarking environment for speech recognition, translation, and summarization models. Unlike typical academic benchmarks like LibriSpeech and FLEURS that uses pre-segmented single-speaker read-speech, FLORAS tests the capabilities of models on raw long-form conversational audio, which can have one or many speakers. To… See the full description on the dataset page: https://huggingface.co/datasets/espnet/floras.audioautomatic-speech-recognition10K<n<100K15 likes3.8k downloads2mo agoHugging Face03espnet /Bagpiper_TTS_SFT_Data Bagpiper-TTS SFT Data Release status: the validated Parquet release is being uploaded. The homepage and metadata may appear before every large shard is committed. Bagpiper-TTS SFT Data supports Bagpiper-TTS, a universal speech-synthesis model that interprets free-form natural-language requests, plans the requested delivery, produces a rich textual caption, and synthesizes the target audio. The release is organized into the six applications used by the paper:… See the full description on the dataset page: https://huggingface.co/datasets/espnet/Bagpiper_TTS_SFT_Data.audiotext-to-speech100K<n<1M0 likes3k downloads2mo agoHugging Face04espnet /Bagpiper_PreTrain_Data Bagpiper Pretraining Data Bagpiper Pretraining Data is the public rich-captioned audio snapshot associated with Bagpiper, an open-ended audio language model that learns bidirectional mappings between audio and comprehensive text descriptions across speech, music, environmental sound, and mixtures. The en metadata describes the primary rich-caption language. Source audio can contain speech or singing in other languages; it is not an English-only audio guarantee. The repository… See the full description on the dataset page: https://huggingface.co/datasets/espnet/Bagpiper_PreTrain_Data.tabularautomatic-speech-recognition10K<n<100K0 likes2.6k downloads2mo agoHugging Face05espnet /ace-opencpop-segments Citation Information @misc{shi2024singingvoicedatascalingup, title={Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing}, author={Jiatong Shi and Yueqian Lin and Xinyi Bai and Keyi Zhang and Yuning Wu and Yuxun Tang and Yifeng Yu and Qin Jin and Shinji Watanabe}, year={2024}, eprint={2401.17619}, archivePrefix={arXiv}, primaryClass={cs.SD}, url={https://arxiv.org/abs/2401.17619}, } audiotext-to-audio100K<n<1M8 likes1k downloads2y agoHugging Face06espnet /mms_ulab_v2MMS ulab v2 is a a massively multilingual speech dataset that contains 8900 hours of unlabeled speech across 4023 languages. In total, it contains 189 language families. It can be used for language identification, spoken language modelling, or speech representation learning. MMS ulab v2 is a reproduced and extended version of the MMS ulab dataset originally proposed in Scaling Speech Technology to 1000+ Languages, covering more languages and containing more data. This dataset includes the raw… See the full description on the dataset page: https://huggingface.co/datasets/espnet/mms_ulab_v2.audioaudio-to-audio10K<n<100K27 likes990 downloads2y agoHugging Face07gptilt /lol-esports-matches GPTilt: League of Legends Esports Matches This dataset is part of the GPTilt open-source initiative, aimed at democratizing access to high-quality LoL data for research and analysis, fostering public exploration, and advancing the community's understanding of League of Legends through data science and AI. It provides a clean, canonical record of the competitive matches and games of professional League of Legends. By using this dataset, users accept full responsibility for any… See the full description on the dataset page: https://huggingface.co/datasets/gptilt/lol-esports-matches.text100K<n<1M0 likes699 downloads12d agoHugging Face08espnet /ace-kising-segments Citation Information @misc{shi2024singingvoicedatascalingup, title={Singing Voice Data Scaling-up: An Introduction to ACE-Opencpop and ACE-KiSing}, author={Jiatong Shi and Yueqian Lin and Xinyi Bai and Keyi Zhang and Yuning Wu and Yuxun Tang and Yifeng Yu and Qin Jin and Shinji Watanabe}, year={2024}, eprint={2401.17619}, archivePrefix={arXiv}, primaryClass={cs.SD}, url={https://arxiv.org/abs/2401.17619}, } audiotext-to-audio10K<n<100K7 likes418 downloads2y agoHugging Face09EsportsBench /EsportsBench EsportsBench: A Collection of Datasets for Benchmarking Rating Systems in Esports EsportsBench is a collection of 20 esports competition datasets. Each row of each dataset represents a match played between either two players or two teams in a professional video game tournament. The goal of the datasets is to provide a resource for comparison and development of rating systems used to predict the results of esports matches based on past results. Date is complete up to 2026-03-31.… See the full description on the dataset page: https://huggingface.co/datasets/EsportsBench/EsportsBench.text1M<n<10M4 likes393 downloads2mo agoHugging Face10espnet /DSUChallenge2024 The Interspeech 2024 Challenge on Speech Processing Using Discrete Units Paper: https://www.isca-archive.org/interspeech_2024/chang24b_interspeech.html Arxiv: https://arxiv.org/abs/2406.07725 Challenge details: https://www.wavlab.org/activities/2024/Interspeech2024-Discrete-Speech-Unit-Challenge/ To cite: @inproceedings{chang24b_interspeech, title = {The Interspeech 2024 Challenge on Speech Processing Using Discrete Units}, author = {Xuankai Chang and Jiatong Shi and… See the full description on the dataset page: https://huggingface.co/datasets/espnet/DSUChallenge2024.audio100K<n<1M1 likes322 downloads2y agoHugging Face11espnet /wikitonguesThe WikiTongues speech corpus is a collection of conversational audio across 700+ languages. It can be used for spoken language modelling or speech representation learning. This dataset includes the raw unsegmented audio in a 16kHz single channel format. Each clip is usually 2-10 minutes long, and contains one or more speakers conversing in their language(s). Sometimes, a speaker may switch languages within a single clip. The total dataset size is around 70 hours. The current version of the… See the full description on the dataset page: https://huggingface.co/datasets/espnet/wikitongues.audioaudio-to-audion<1K4 likes256 downloads2y agoHugging Face12gptilt /lol-esports-entities GPTilt: League of Legends Esports Directory This dataset is part of the GPTilt open-source initiative, aimed at democratizing access to high-quality LoL data for research and analysis, fostering public exploration, and advancing the community's understanding of League of Legends through data science and AI. It provides a clean, canonical reference for the people and organizations of competitive League of Legends. By using this dataset, users accept full responsibility for any… See the full description on the dataset page: https://huggingface.co/datasets/gptilt/lol-esports-entities.text100K<n<1M1 likes249 downloads12d agoHugging Face13espnet /ml_superb_hfaudio100K<n<1M7 likes230 downloads2y agoHugging Face14espejelomar /go2-air-controlbench-v1 Go2 Air ControlBench v1 Public Preview Go2 Air ControlBench is a compact command-to-outcome benchmark for a stock Unitree Go2 Air. It asks a narrow question that matters for robot planners and world-model scorers: If we command the robot to move, what actually happens, and which candidate command should a planner have selected? The release is deliberately small and claim-bounded. It is not an imitation-learning corpus and not mocap-grade ground truth. It is a public-safe… See the full description on the dataset page: https://huggingface.co/datasets/espejelomar/go2-air-controlbench-v1.imagerobotics1K<n<10K0 likes218 downloads3mo agoHugging Face15espejelomar /so101-can-butler SO-101 Can Butler Teleoperated demonstrations of a human-triggered can handover on a low-cost SO-101 arm: the robot stays still until a person's hand appears on the mat, then reaches, grasps a can, and hands it over. A SmolVLA fine-tune on this data — 20k steps, about $2.30 of rented RTX 4090 — grasped and delivered the can autonomously, verified 3 times out of 3 attempts. Model: espejelomar/smolvla-so101-can-butler. What it looks like A policy trained on… See the full description on the dataset page: https://huggingface.co/datasets/espejelomar/so101-can-butler.tabularrobotics10K<n<100K0 likes202 downloads26d agoHugging Face16lab260 /espeech_balalaika ESpeech datasets (w/o podcasts) Annotated by Balalaika [!IMPORTANT] Official dataset for our INTERSPEECH 2026 paper "A Data-Centric Framework for Addressing Phonetic and Prosodic Challenges in Russian Speech Generative Models" (arXiv:2507.13563). Part of the Balalaika Russian speech data-processing pipeline — code: https://github.com/lab260ru/balalaika. If you use this resource, please cite it. A curated Russian speech dataset for advanced speech generative tasks.… See the full description on the dataset page: https://huggingface.co/datasets/lab260/espeech_balalaika.tabulartext-to-speech100K<n<1M4 likes186 downloads3mo agoHugging Face17hieu24 /esp32-arm-test2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "base.pos", "shoulder.pos", "elbow.pos", "wrist.pos", "gripper.pos" ], "shape": [ 5 ] }, "observation.state": {… See the full description on the dataset page: https://huggingface.co/datasets/hieu24/esp32-arm-test2.tabularrobotics1K<n<10K0 likes155 downloads17d agoHugging Face18MiguelGP-13 /JobOffers_ESP Ofertas de Empleo Públicas en España (EURES, 2025) https://doi.org/10.57967/hf/6740 Dataset Summary Este dataset recopila ofertas de empleo publicadas en portales oficiales de empleo europeos y españoles, principalmente a través de la red EURES (European Employment Services).Forma parte del proyecto desarrollado para una práctica en la asignatura Descubrimiento del Conocimiento en Datos Complejos del grado de Ciencia de Datos e Inteligencia Artificial en la Universidad… See the full description on the dataset page: https://huggingface.co/datasets/MiguelGP-13/JobOffers_ESP.texttext-classification1K<n<10K2 likes151 downloads11mo agoHugging Face19espnet /jesus_dramasJesus Dramas is a collection of religious audio dramas across 430 languages. In total, there is around 640 hours of audio. It can be used for language identification, spoken language modelling, or speech representation learning. This dataset includes the raw unsegmented audio in a 16kHz single channel format. Each audio drama can have multiple speakers, for both male and female voices. It can be segmented into utterances with a voice activity detection (VAD) model such as this one. The… See the full description on the dataset page: https://huggingface.co/datasets/espnet/jesus_dramas.audioaudio-to-audion<1K4 likes137 downloads2y agoHugging Face20jensjepsen /esperanto-mt-parallel-v13 esperanto-mt-parallel v13 EN<->EO parallel training corpus. 5,025,333 rows after dedup. What changed vs v12 (jensjepsen/esperanto-mt-parallel) Dropped Helsinki-NLP/opus-100 en-eo train split (144,549 rows). It is an aggregated multilingual blob that bundles KDE4/GNOME/Ubuntu .po localization pairs without src labels. In v12 this caused a systematic MT failure mode: capitalized-fragment-no-terminal-punct inputs collapsed to memorized UI labels (e.g. "@ info:… See the full description on the dataset page: https://huggingface.co/datasets/jensjepsen/esperanto-mt-parallel-v13.texttranslation1M<n<10M0 likes125 downloads2mo agoHugging Face21espnet /long-yodas-unsegmentedaudio1K<n<10K0 likes118 downloads2y agoHugging Face22espejelomar /worldforge-go2-dimos-replay-world-pairs WorldForge Go2 DimOS Replay World Pairs This dataset is a compact, derived world-model dataset built from public dimensionalOS/dimos Unitree Go2 replay assets. Companion benchmark: go2-air-controlbench-v1 provides measured command-to-outcome trials on a real Go2 (commands, no images). This dataset provides the robot-POV image pairs (images, no commands). Together they cover the visual and control halves of the WorldForge score workflow. New — expanded config:… See the full description on the dataset page: https://huggingface.co/datasets/espejelomar/worldforge-go2-dimos-replay-world-pairs.imagerobotics10K<n<100K2 likes113 downloads3mo agoHugging Face23J-joon /aloha_sim_insertion_espada_finalThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.0", "robot_type": "aloha", "total_episodes": 50, "total_frames": 20000, "total_tasks": 1, "total_videos": 50, "total_chunks": 1, "chunks_size": 1000, "fps": 50, "splits": { "train": "0:50" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/J-joon/aloha_sim_insertion_espada_final.tabularrobotics10K<n<100K0 likes107 downloads1y agoHugging Face24robots123 /yam-espressoThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "yam_follower", "total_episodes": 91, "total_frames": 72048, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:91" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/robots123/yam-espresso.tabularrobotics10K<n<100K0 likes97 downloads9mo agoHugging Face25Teklia /Esposalles-line Esposalles - line level Dataset Summary The Marriage Licenses ground-truth is compiled from the Marriage Licenses Books conserved at the Archives of the Cathedral of Barcelona. Note that all images are resized to a fixed height of 128 pixels. Languages All the documents in the dataset are written in Catalan. Dataset Structure Data Instances { 'image': <PIL.JpegImagePlugin.JpegImageFile image mode=RGB size=1244x128 at 0x1A800E8E190… See the full description on the dataset page: https://huggingface.co/datasets/Teklia/Esposalles-line.imageimage-to-text1K<n<10K1 likes82 downloads3y agoHugging Face26cthorrez /EsportsBenchTestTESTING EsportsBench: A Collection of Datasets for Benchmarking Rating Systems in Esports EsportsBench is a collection of 20 esports competition datasets. Each row of each dataset represents a match played between either two players or two teams in a professional video game tournament. The goal of the datasets is to provide a resource for comparison and development of rating systems used to predict the results of esports matches based on past results. Date is complete up to… See the full description on the dataset page: https://huggingface.co/datasets/cthorrez/EsportsBenchTest.text1M<n<10M0 likes77 downloads2mo agoHugging Face27yiyic /culturaX_esptext100K<n<1M2 likes68 downloads2y agoHugging Face28espada105 /augmented-brain-tumor-segmentation-v2image10K<n<100K0 likes63 downloads1y agoHugging Face29ia-espirita /pinga-fogo-chico-xavier 🎙️ Pinga-Fogo com Chico Xavier — TV Tupi, 1971 As duas entrevistas históricas do médium Chico Xavier, transmitidas ao vivo pela TV Tupi em 1971, transcritas e estruturadas em turnos de fala com timestamp. 345 turnos (115 deles respostas do próprio Chico Xavier), a partir de 6 horas de áudio — o registro mais extenso do médium falando de improviso, sem edição, diante de um painel de jornalistas. Arquivos Arquivo Programa Turnos Respostas do Chico… See the full description on the dataset page: https://huggingface.co/datasets/ia-espirita/pinga-fogo-chico-xavier.tabularquestion-answeringn<1K1 likes63 downloads1mo agoHugging Face30J-joon /aloha_sim_transfer_cube_espadaThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.0", "robot_type": "aloha", "total_episodes": 50, "total_frames": 6600, "total_tasks": 1, "total_videos": 50, "total_chunks": 1, "chunks_size": 1000, "fps": 50, "splits": { "train": "0:50" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/J-joon/aloha_sim_transfer_cube_espada.tabularrobotics1K<n<10K0 likes62 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.