datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DuplexConv
DuplexConv
DuplexConv is a large-scale Chinese multi-channel conversational speech dataset with LLM-assisted annotations, developed by ASLP@NPU and QualiaLabs as part of the SmoothConv–DuplexConv corpus family.
Companion dataset: SmoothConv on HuggingFace (100 hours, expert human annotation). DuplexConv and SmoothConv share the same conversational domains and a unified data design. SmoothConv focuses on high-quality human annotations for benchmarking and… See the full description on the dataset page: https://huggingface.co/datasets/qualialabsAI/DuplexConv.cmd-audio-dumpimage-manipulation-dataset-compilationGelbooru-Post-DumpUni10K
Uni10K from
A LoD of Gaussians: Unified Training and Rendering for Ultra-Large-Scale Reconstruction with External Memory
Felix Windisch1
Thomas Köhler1
Lukas Radl1
Mattia D'Urso1
Michael Steiner1
Dieter Schmalstieg1,2
Markus Steinberger1,3
1Graz University of Technology,
2University of Stuttgart,
3Huawei Technologies… See the full description on the dataset page: https://huggingface.co/datasets/mattia-durso/Uni10K.datasetsllava-odanime_dump_redditviet-tts-datasedurg-university-noticesravdessmsmarco-passage-v2-duplicate-idsosu-beatmaps-duplicated
osu! Beatmaps Dataset (WebDataset)
A collection of ranked/loved osu! beatmaps with audio and chart data, in WebDataset format.
Dataset Variants
Variant
Audio Format
Description
original
MP3/OGG/WAV
Full quality original audio files
compressed
64kbps Mono Opus
Compressed audio for smaller download
from datasets import load_dataset
# Load original audio variant
ds = load_dataset("project-riz/osu-beatmaps", "original", streaming=True)
# Load compressed… See the full description on the dataset page: https://huggingface.co/datasets/IamXiangyu/osu-beatmaps-duplicated.otoSpeech-HQ-full-duplex-samples
Dataset Card for otoSpeech-HQ-full-duplex-samples: Full-Duplex Conversational Speech Dataset Samples
Dataset Summary
otoSpeech-HQ-full-duplex-samples is a curated collection of high-quality full-duplex conversational speech samples designed for commercial and production-oriented use.
This repository is derived from a private subset of otoSpeech and features carefully selected English two-speaker conversations with enhanced audio quality. The samples are intended for… See the full description on the dataset page: https://huggingface.co/datasets/otoearth/otoSpeech-HQ-full-duplex-samples.PMC_OA_Dental
Dental PMC Image-Caption Dataset
数据集描述
这是一个口腔医学领域的图像-标题对数据集,来源于PubMed Central (PMC) 开放获取文章。
数据集内容
图片数量: 22张医学图片
格式: JPG和PNG(GIF转换)
每张图片包含:
图像数据(jpg或png格式)
标题/描述文本(txt)
丰富的元数据(json)
元数据字段
{
"accession_id": "PMC文章ID",
"pmid": "PubMed ID",
"title": "文章标题",
"journal": "期刊名称",
"image_id": "图片ID",
"image_file_name": "原始文件名",
"image_hash": "图片哈希值",
"license": "许可证(通常为cc-by)",
"keywords": ["关键词列表"],
"context": ["相关上下文段落"]
}
数据来源… See the full description on the dataset page: https://huggingface.co/datasets/duanyt/PMC_OA_Dental.nans_tab_dump_1234lj_speech_1.1misl-image-db-70c-wdsHOMED_dataDual-VLA-10Hz
