tone
Datasets
All datasets matching “tone”TST100K
TST100K: A Large-Scale Triplet Dataset for Tone Style Transfer
Project Page | Paper
TST100K is a large-scale dataset for reference-based tone style transfer in photo retouching. The current release contains 108,808 image triplets constructed from the public PPR10K, MIT-Adobe FiveK, and Food-101 research datasets.
Each triplet contains a content image, a reference image, and a ground-truth target image. The task is to reproduce the photographic color and tone of the reference… See the full description on the dataset page: https://huggingface.co/datasets/ToneStyle/TST100K.TST2K
TST2K: A High-Quality Benchmark for Tone Style Transfer
TST2K is a benchmark for reference-based tone style transfer in photo retouching. It contains 2,000 image triplets built from the public PPR10K and MIT-Adobe FiveK research datasets.
Each triplet includes a content image, a reference image, and a ground-truth target image. The goal is to match the color and tone of the reference image while keeping the content and structure of the original image. TST2K covers a range of… See the full description on the dataset page: https://huggingface.co/datasets/ToneStyle/TST2K.ToneWebinars
ToneWebinars
ToneWebinars — это обработанная версия ZeroAgency/shkolkovo-bobr.video-webinars-audio.
Оригинальные файлы были перепаковыны в parquet формат с нарезкой по приложенным такмкодам. В датасете 2053.55 часа аудио для train сплита и 154.34 для validation.
Описание
Для каждого примера прдоставляются:
Ссылка на MP3-файл (audio)
Текстовая расшифровка (text)
Частота дискретизации (sample_rate)
Формат записи (JSON)
{
"audio":… See the full description on the dataset page: https://huggingface.co/datasets/Vikhrmodels/ToneWebinars.ToneBooksPlus
ToneBooksPlus
ToneBooksPlus — расширенная версия датасета Vikhrmodels/ToneBooks, но без эмоциональной разметки. В датасете 179.16 часов аудио для train сплита и 9.42 часа для validation.
Большое спасибо its5Q за помощь в сборе этих данных.
Описание
Для каждого аудиофрагмента собраны:
Ссылка на MP3-файл (audio)
Текстовая расшифровка (text)
Имя голоса (voice_name) — одно из имён дикторов:
Aleksandr Kotov
Aleksandr Zbarovskii
Alina Archibasova
Daniel Che… See the full description on the dataset page: https://huggingface.co/datasets/Vikhrmodels/ToneBooksPlus.ToneSlavic
ToneSlavic
ToneSlavic — это датасет, собранный из русских, белорусских и украинских сплитов датасета Common Voice Corpus 21.0. В выборку вошли только фразы из файлов validated.tsv.
Описание
Для каждого примера предоставляются:
Ссылка на MP3-файл (audio)
Текстовая расшифровка (text)
Язык примера (locale)
ru — русский
be — белорусский
uk — украинский
В датасете представлены следующие статистики по сплитам и языкам:
Train (1 475 164 строк, 1 968.22 часов аудио):… See the full description on the dataset page: https://huggingface.co/datasets/Vikhrmodels/ToneSlavic.ToneWebinars
ToneWebinars

