datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dl3dv_bench_torch_960dl3dv_mixMode_1Sub2D_TrueAlpha_0.95MixWeight_unionMask_FalseDetach3DDL3DV-Evaluation
DL3DV Testing Split Download Instructions
This repo contains all 55 scenes for evaluation. Note: it is an independent dataset, and none of its scenes overlap with those in DL3DV-10K. Have a galance on the preview page: https://dl3dv-10k.github.io/DL3DV-Testing-Split-Preview/.
Download
As the whole benchmark dataset is ~500G, a python script to download and untar files.
Environment Setup
The download script relies on huggingface hub, tqdm. You can download by… See the full description on the dataset page: https://huggingface.co/datasets/DL3DV/DL3DV-Evaluation.mvsplat_dl3dv_2dMode_1Sub2D_FalseAlphadl3dv_mixMode_1Sub2D_TrueAlpha_0.95MixWeight_unionMask_TrueDetach3Dtransplat_dl3dv_2dMode_1Sub2D_FalseAlphadl3dv_2dMode_1Sub2D_TrueAlpha_fromRe10KPretraineddl3dv_InP_480dl3dvDL3DV-2k
DL3DV-2K
📖Paper
| 🏠Homepage
| 🤗ETCHR-FLUX.2-klein-9B Model
| 🤗ETCHR SFT-400K Dataset
| 🤗ETCHR GRPO-10K Dataset
| 🤗DL3DV-2K Benchmark
DL3DV-2K is a benchmark constructed from the DL3DV dataset for evaluating the viewpoint transformation capability of large models in spatial reasoning tasks, comprising 2K samples in total. Each sample contains: images (original images), aux_images (transformed images provided for human reference only and not used as question input)… See the full description on the dataset page: https://huggingface.co/datasets/internlm/DL3DV-2k.dl3dv_960p_testtransplat_dl3dv_2dMode_1Sub2D_TrueAlphadl3dv_960_processed_captions_qwen3_32bdl3dv_jsondl3dv_2dMode_1Sub2D_FalseAlphadl3dv_2dMode_1Sub2D_TrueAlphaDL3DV-GSdl3dv_mixMode_1Sub2D_TrueAlpha_fromRe10KPretrained_0.95MixWeight_unionMask_TrueDetach3Ddl3dv_InP_480_latentmvsplat_dl3dv_2dMode_1Sub2D_TrueAlphamvsplat_dl3dv_mixMode_1Sub2D_TrueAlpha_0.95MixWeight_unionMask_TrueDetach3D
