yue
Datasets
All datasets matching “yue”Objaverse-PBR-render
Objaverse-PBR-render
PBR rendered condition videos for Ink3D — a 3D texture generation pipeline using video generative models.
📄 Paper: Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models💻 Code: github.com/YueHan99/Ink3D.TextureGen🤗 Model: Yuehavingfun/orbitpainter-single
This dataset provides pre-rendered geometry condition videos (position, normal, albedo, RGB, depth, mask) for ~23,000 Objaverse models, rendered from horizontal (H) and… See the full description on the dataset page: https://huggingface.co/datasets/Yuehavingfun/Objaverse-PBR-render.regx-benchmark
RegX
Cross-Domain Multi-View Point Cloud Registration Benchmark
RegX evaluates multi-view point cloud registration across scales spanning nine orders
of magnitude — nanometre-scale microscopy to kilometre-scale airborne maps — and sensors
never designed to be compared: clinical colonoscopes, RGB-D cameras, spinning and
solid-state LiDAR, terrestrial and airborne laser scanners.
Most registration benchmarks fix one sensor and one scale. RegX asks a narrower question
instead: does… See the full description on the dataset page: https://huggingface.co/datasets/YuePanEdward/regx-benchmark.Movie101
Movie101
[!NOTE]
Please carefully read the Movie101 license before using the data.Current dataset version: Movie101v2
Audio Description (AD) describes movie content in real time to help visually impaired individuals enjoy movies, where a narration speech briefly summarizes the ongoing plots during pauses in character dialogue, help its audience keep up with the movie.
The AD creation involves extensive work by human experts, which is costly and difficult to cover the vast array… See the full description on the dataset page: https://huggingface.co/datasets/yuezih/Movie101.youtube_caption_yue
YouTube ASR Caption Dataset (Cantonese)
This dataset was built from YouTube videos with manually provided captions in Cantonese. We used SenseVoice to re-transcribe the audio and filtered segments to build a high-quality collection of audio-caption pairs.
What’s included
Segments where the ASR output is identical to the original caption — likely clean.
Segments where differences are only homophones (同音字) or English words — likely ASR mistakes.
This combination supports… See the full description on the dataset page: https://huggingface.co/datasets/ming030890/youtube_caption_yue.yue2-minted-corpus
YuE2 minted corpus
Songs generated by m-a-p/YuE2-3B (with YuE2-Vae) from
MusicForge plan_50k prompts, queued genre-round-robin (355 genres, ~35% instrumental), seeds from the plan. Every track keeps the exact
intermediate the model produced, so the set is a labelled corpus for training an audio → semantic-token encoder (the piece YuE2 does not ship)
and a grammar regularizer for AR fine-tunes.
tracks/<pid>/
file
content
audio.flac
48 kHz stereo 24-bit
semantic.npy… See the full description on the dataset page: https://huggingface.co/datasets/Mothersuperior/yue2-minted-corpus.my-cloudpaste-storage
