datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
scannetppHarmonizer-Dataset
HARMONIZER DATASET
Dataset Description
Training dataset for DiffusionHarmonizer: a generative AI model for image and video enhancement bridging neural reconstruction and photorealistic simulation .
Model checkpoints: https://huggingface.co/nvidia/Harmonizer/Training code: https://github.com/NVIDIA/harmonizer/
The dataset was curated to support the following functions of the model:
3D reconstruction artifact removal
Harmonization of inserted objects to blend… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Harmonizer-Dataset.blendedmvstiered_imagenetwildrgbdwaymoadvisor-twinsNot for commercial use. These are only being used to showcase Gaia as a framework to build personalised AI agents.
AVQA-R1-6KThis repository contains data presented in EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning.
For training and inference, please refer to the Code: https://github.com/HarryHsing/EchoInk
Data Format in AVQA-R1-6K:
{
"problem_id": 0,
"problem": "What is the source of the sound in the video?",
"data_type": "image_audio",
"problem_type": "multiple choice",
"options": [
"A. motorcycle",
"B. automobile"… See the full description on the dataset page: https://huggingface.co/datasets/harryhsing/AVQA-R1-6K.gaia-whitepaperhard_math_wavwormhole-docssolidityboundlesscardano_cipsMulti-turn-editingsuper_agent_advisornanowmyoga_poses_collectionno-cbdinfinigenmetamask-snapspeter-thielharish_dicmvs_synthtsfm-har-benchvana-redditdaoboardroom_api_demosteve-jobs
