CoolFace
22 results

tone

ToneStyle /TST100K TST100K: A Large-Scale Triplet Dataset for Tone Style Transfer Project Page | Paper TST100K is a large-scale dataset for reference-based tone style transfer in photo retouching. The current release contains 108,808 image triplets constructed from the public PPR10K, MIT-Adobe FiveK, and Food-101 research datasets. Each triplet contains a content image, a reference image, and a ground-truth target image. The task is to reproduce the photographic color and tone of the reference… See the full description on the dataset page: https://huggingface.co/datasets/ToneStyle/TST100K.image100K<n<1M2 likes3.4k downloads2mo agoHugging FaceToneStyle /TST2K TST2K: A High-Quality Benchmark for Tone Style Transfer TST2K is a benchmark for reference-based tone style transfer in photo retouching. It contains 2,000 image triplets built from the public PPR10K and MIT-Adobe FiveK research datasets. Each triplet includes a content image, a reference image, and a ground-truth target image. The goal is to match the color and tone of the reference image while keeping the content and structure of the original image. TST2K covers a range of… See the full description on the dataset page: https://huggingface.co/datasets/ToneStyle/TST2K.imageimage-to-image1K<n<10K1 likes1.1k downloads3mo agoHugging FaceVikhrmodels /ToneWebinars ToneWebinars ToneWebinars — это обработанная версия ZeroAgency/shkolkovo-bobr.video-webinars-audio. Оригинальные файлы были перепаковыны в parquet формат с нарезкой по приложенным такмкодам. В датасете 2053.55 часа аудио для train сплита и 154.34 для validation. Описание Для каждого примера прдоставляются: Ссылка на MP3-файл (audio) Текстовая расшифровка (text) Частота дискретизации (sample_rate) Формат записи (JSON) { "audio":… See the full description on the dataset page: https://huggingface.co/datasets/Vikhrmodels/ToneWebinars.audioautomatic-speech-recognition100K<n<1M13 likes771 downloads1y agoHugging FaceVikhrmodels /ToneBooksPlus ToneBooksPlus ToneBooksPlus — расширенная версия датасета Vikhrmodels/ToneBooks, но без эмоциональной разметки. В датасете 179.16 часов аудио для train сплита и 9.42 часа для validation. Большое спасибо its5Q за помощь в сборе этих данных. Описание Для каждого аудиофрагмента собраны: Ссылка на MP3-файл (audio) Текстовая расшифровка (text) Имя голоса (voice_name) — одно из имён дикторов: Aleksandr Kotov Aleksandr Zbarovskii Alina Archibasova Daniel Che… See the full description on the dataset page: https://huggingface.co/datasets/Vikhrmodels/ToneBooksPlus.audiotext-to-speech100K<n<1M7 likes468 downloads1y agoHugging FaceVikhrmodels /ToneSlavic ToneSlavic ToneSlavic — это датасет, собранный из русских, белорусских и украинских сплитов датасета Common Voice Corpus 21.0. В выборку вошли только фразы из файлов validated.tsv. Описание Для каждого примера предоставляются: Ссылка на MP3-файл (audio) Текстовая расшифровка (text) Язык примера (locale) ru — русский be — белорусский uk — украинский В датасете представлены следующие статистики по сплитам и языкам: Train (1 475 164 строк, 1 968.22 часов аудио):… See the full description on the dataset page: https://huggingface.co/datasets/Vikhrmodels/ToneSlavic.audioautomatic-speech-recognition1M<n<10M0 likes300 downloads1y agoHugging FaceICML-2026 /ToneWebinars ToneWebinars audioautomatic-speech-recognition100K<n<1M0 likes278 downloads8mo agoHugging Face

People