CoolFace
20 results

yue

Yuehavingfun /Objaverse-PBR-render Objaverse-PBR-render PBR rendered condition videos for Ink3D — a 3D texture generation pipeline using video generative models. 📄 Paper: Ink3D: Sculpting 3D Assets with Extremely Complex Textures via Video Generative Models💻 Code: github.com/YueHan99/Ink3D.TextureGen🤗 Model: Yuehavingfun/orbitpainter-single This dataset provides pre-rendered geometry condition videos (position, normal, albedo, RGB, depth, mask) for ~23,000 Objaverse models, rendered from horizontal (H) and… See the full description on the dataset page: https://huggingface.co/datasets/Yuehavingfun/Objaverse-PBR-render.image-to-video0 likes11k downloads3mo agoHugging FaceYuePanEdward /regx-benchmark RegX Cross-Domain Multi-View Point Cloud Registration Benchmark RegX evaluates multi-view point cloud registration across scales spanning nine orders of magnitude — nanometre-scale microscopy to kilometre-scale airborne maps — and sensors never designed to be compared: clinical colonoscopes, RGB-D cameras, spinning and solid-state LiDAR, terrestrial and airborne laser scanners. Most registration benchmarks fix one sensor and one scale. RegX asks a narrower question instead: does… See the full description on the dataset page: https://huggingface.co/datasets/YuePanEdward/regx-benchmark.3dother1K<n<10K2 likes5.3k downloads19d agoHugging Faceyuezih /Movie101gated Movie101 [!NOTE] Please carefully read the Movie101 license before using the data.Current dataset version: Movie101v2 Audio Description (AD) describes movie content in real time to help visually impaired individuals enjoy movies, where a narration speech briefly summarizes the ongoing plots during pauses in character dialogue, help its audience keep up with the movie. The AD creation involves extensive work by human experts, which is costly and difficult to cover the vast array… See the full description on the dataset page: https://huggingface.co/datasets/yuezih/Movie101.imagevideo-text-to-text100K<n<1M7 likes4.3k downloads1y agoHugging Faceming030890 /youtube_caption_yue YouTube ASR Caption Dataset (Cantonese) This dataset was built from YouTube videos with manually provided captions in Cantonese. We used SenseVoice to re-transcribe the audio and filtered segments to build a high-quality collection of audio-caption pairs. What’s included Segments where the ASR output is identical to the original caption — likely clean. Segments where differences are only homophones (同音字) or English words — likely ASR mistakes. This combination supports… See the full description on the dataset page: https://huggingface.co/datasets/ming030890/youtube_caption_yue.audio10K<n<100K2 likes3.8k downloads1y agoHugging FaceMothersuperior /yue2-minted-corpus YuE2 minted corpus Songs generated by m-a-p/YuE2-3B (with YuE2-Vae) from MusicForge plan_50k prompts, queued genre-round-robin (355 genres, ~35% instrumental), seeds from the plan. Every track keeps the exact intermediate the model produced, so the set is a labelled corpus for training an audio → semantic-token encoder (the piece YuE2 does not ship) and a grammar regularizer for AR fine-tunes. tracks/<pid>/ file content audio.flac 48 kHz stereo 24-bit semantic.npy… See the full description on the dataset page: https://huggingface.co/datasets/Mothersuperior/yue2-minted-corpus.audio-to-audio5 likes3.2k downloads9d agoHugging Faceyuezhengling /my-cloudpaste-storageimagen<1K0 likes1.2k downloads7mo agoHugging Face