datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
pianos
Dataset Card for Piano Sound Quality Dataset
The original dataset is sourced from the Piano Sound Quality Dataset, which includes 12 full-range audio files in .wav/.mp3/.m4a format representing seven models of pianos: Kawai upright piano, Kawai grand piano, Young Change upright piano, Hsinghai upright piano, Grand Theatre Steinway piano, Steinway grand piano, and Pearl River upright piano. Additionally, there are 1,320 split monophonic audio files in .wav/.mp3/.m4a format, bringing… See the full description on the dataset page: https://huggingface.co/datasets/ccmusic-database/pianos.music_genre
Dataset Card for Music Genre
The Default dataset comprises approximately 1,700 musical pieces in .mp3 format, sourced from the NetEase music. The lengths of these pieces range from 270 to 300 seconds. All are sampled at the rate of 22,050 Hz. As the website providing the audio music includes style labels for the downloaded music, there are no specific annotators involved. Validation is achieved concurrently with the downloading process. They are categorized into a total of 16… See the full description on the dataset page: https://huggingface.co/datasets/ccmusic-database/music_genre.bel_canto
Dataset Card for Bel Conto and Chinese Folk Song Singing Tech
Original Content
This dataset is created by the authors and encompasses two distinct singing styles: bel canto and Chinese folk singing. Bel canto is a vocal technique frequently employed in Western classical music and opera, symbolizing the zenith of vocal artistry within the broader Western musical heritage. Chinese folk singing, for which there is no official English translation, is referred to here as a… See the full description on the dataset page: https://huggingface.co/datasets/ccmusic-database/bel_canto.Traffic_Sign_Recogntion_DatabaseTSRD (Traffic Sign Recognition Database) 是一个中国交通标志数据集,包含多种交通标志类别。数据集分为训练集和测试集:
训练集:包含约4170张图像
测试集:包含约1994张图像
类别数:约58个不同的交通标志类别
数据集格式为:
图像文件名;宽;高;x1;y1;x2;y2;类别;
包含全种类数据集 / 4方向指示牌数据集
danbooru2023-metadata-database
Metadata Database for Danbooru2023
Danbooru 2023 datasets: https://huggingface.co/datasets/nyanko7/danbooru2023
The latest entry of this database is id 7,866,491. Which is newer than nyanko7's dataset.
This dataset contains a sqlite db file which have all the tags and posts metadata in it.
The Peewee ORM config file is provided too, plz check it for more information. (Especially on how I link posts and tags together)
The original data is from the official dump of the posts info.… See the full description on the dataset page: https://huggingface.co/datasets/KBlueLeaf/danbooru2023-metadata-database.ArASL_Database_Grayscale
Dataset Card for "ArASL_Database_Grayscale"
Dataset Summary
A new dataset consists of 54,049 images of ArSL alphabets performed by more than 40 people for 32 standard Arabic signs and alphabets.
The number of images per class differs from one class to another. Sample image of all Arabic Language Signs is also attached. The CSV file contains the Label of each corresponding Arabic Sign Language Image based on the image file name.
Supported Tasks and Leaderboards… See the full description on the dataset page: https://huggingface.co/datasets/pain/ArASL_Database_Grayscale.ourdream-database
OurDream Character Database
Complete database of 20,884 AI characters from ourdream.ai, categorized as Women (12,492) and Trans (8,392).
Dataset Structure
database/all_characters.json — Full character database with 20,884 entries
Each entry includes: id, displayId, name, gender, style, age, likeCount, messageCount, tags, shortDescription, thumbUrl, category
Categories
Category
Count
Women
12,492
Trans
8,392
Total
20,884… See the full description on the dataset page: https://huggingface.co/datasets/lcuifer0/ourdream-database.
