tatar
Datasets
All datasets matching “tatar”tatartts_vibevoice-formattatara_kogasa_touhou
Dataset of tatara_kogasa/多々良小傘/타타라코가사 (Touhou)
This is the dataset of tatara_kogasa/多々良小傘/타타라코가사 (Touhou), containing 500 images and their tags.
The core tags of this character are blue_hair, short_hair, red_eyes, blue_eyes, heterochromia, which are pruned in this dataset.
Images are crawled from many sites (e.g. danbooru, pixiv, zerochan ...), the auto-crawling system is powered by DeepGHS Team(huggingface organization).
List of Packages
Name
Images
Size… See the full description on the dataset page: https://huggingface.co/datasets/CyberHarem/tatara_kogasa_touhou.TatarTTS
TatarTTS Dataset
Paper: TatarTTS: An Open-Source Text-to-Speech Synthesis Dataset for the Tatar Language
GitHub: https://github.com/IS2AI/TatarTTS
Description: TatarTTS is an open-source text-to-speech dataset for the Tatar language.
The dataset comprises ~70 hours of transcribed audio recordings, featuring two professional speakers (one male and one female).
Citation:
The project was developed in academic collaboration between ISSAI and Institute of Applied Semiotics of Tatarstan… See the full description on the dataset page: https://huggingface.co/datasets/issai/TatarTTS.sampled-tatar-datasetSampled Tatar dataset based on https://huggingface.co/datasets/HuggingFaceFW/fineweb-2
tatar-speech-commands
An Open-Source Tatar Speech Commands Dataset
Paper: Paper
An Open-Source Tatar Speech Commands Dataset for IoT and Robotics Applications
GitHub: https://github.com/IS2AI/TatarSCR
Description:
The dataset covers 35 commands used in robotics, IoT, and smart systems. In total, the dataset contains 3,547 one-second utterances from 153 people. The utterances were saved in the WAV format with a sampling rate of 16 kHz.
Citation: The project was developed in academic collaboration between… See the full description on the dataset page: https://huggingface.co/datasets/issai/tatar-speech-commands.tatar-russian-parallel-corporaТатарско-русский параллельный корпус.
@inproceedings{
title={Tatar parallel corpus},
author={Academy of Siences of the Recpublic of Tatarstan, Institute of Applied Semiotics.},
year={2023}
}
