CoolFace
8 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01XRXRX /X-Voice-TestsetX-Voice Multilingual Test Set High-Fidelity Test Set for Multilingual Text-to-Speech across 30 Languages This test set is built as part of the research: X-Voice: One Speaker, 30+ Languages with Zero-Shot Voice Cloning, serving as the evaluation benchmark for our model. Dataset Summary 30 languages European: bg (Bulgarian), cs (Czech), da (Danish), de (German), el (Greek), en (English), es (Spanish), et (Estonian), fi (Finnish), fr (French), hr (Croatian), hu (Hungarian), it… See the full description on the dataset page: https://huggingface.co/datasets/XRXRX/X-Voice-Testset.audiotext-to-speech10K<n<100K4 likes72 downloads5mo agoHugging Face02KoddaDuck /testsplitaudion<1K1 likes10 downloads3y agoHugging Face03omkarthawakar /test-setimage1K<n<10K0 likes8 downloads1y agoHugging Face04marian-nmt /marian-regression-tests Marian Regression Tests This repository contains datasets and models for marian regression tests. The regression test suite is at https://github.com/marian-nmt/marian-regression-tests textn<1K0 likes6 downloads2y agoHugging Face05Yougen /asc_testsetgated Yougen/asc_testset Audio Scene Classification (ASC) speech dataset, packed as WebDataset tar shards. Layout data/ train/ metadata.csv audio/ train-000.tar train-001.tar ... validation/ metadata.csv audio/ validation-000.tar ... test/ metadata.csv audio/ test-000.tar ... Shard counts: test_a1: 8 tar shard(s) test_a2: 16 tar shard(s) test_a3: 13 tar shard(s) test_a4: 15 tar shard(s) test_a5: 8 tar… See the full description on the dataset page: https://huggingface.co/datasets/Yougen/asc_testset.audioaudio-classification100K<n<1M0 likes6 downloads5mo agoHugging Face06xxxllz /TEST-SPECTRAtextn<1K0 likes5 downloads5mo agoHugging Face07HGYash /test_shardsimagen<1K0 likes3 downloads10mo agoHugging Face08Yougen /said_testsetgated Yougen/said_testset Speaker Recognition (SRE) speech dataset, packed as WebDataset tar shards. Layout data/ train/ metadata.csv audio/ train-000.tar train-001.tar ... validation/ metadata.csv audio/ validation-000.tar ... test/ metadata.csv audio/ test-000.tar ... Shard counts: test: 20 tar shard(s) Inside each tar, every sample is a pair sharing a unique key: <key>.wav # raw audio bytes… See the full description on the dataset page: https://huggingface.co/datasets/Yougen/said_testset.audioaudio-classification10K<n<100K0 likes1 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.