visualization
visualization-chartsWildDet3D-visualization-source
WildDet3D Visualization Data
This repository hosts the visualization data for the WildDet3D-Bench benchmark — a human-annotated evaluation set for monocular 3D object detection in the wild.
Dataset Overview
WildDet3D-Bench is a validation set of 2,470 images drawn from three source datasets, with 9,256 human-verified 3D bounding box annotations across 2,196 images.
Source
Images
Description
COCO Val
424
MS-COCO 2017 validation
LVIS Train
1,113
LVIS v1.0 (COCO… See the full description on the dataset page: https://huggingface.co/datasets/allenai/WildDet3D-visualization-source.itw_pipeline_visualization
ITW Pipeline Visualization — assets
Purpose of this repository
This repository exists for one reason: to serve media files to the
visualization site at
https://silicon23.github.io/itw_pipeline_visualization/.
A static site cannot host its own heavy media, so the frames, depth maps and
overlay renders it streams live here.
It is not published for redistribution, and it is not a dataset to train or
evaluate on. It is the asset backing of a figure — the equivalent… See the full description on the dataset page: https://huggingface.co/datasets/Silicon23/itw_pipeline_visualization.visualizationSOC-Training-Data-Visualization
Paper Link
SOS: Synthetic Object Segments Improve Detection, Segmentation, and Grounding
Code repo
Code for Generation
Citation
@misc{huang2025sossyntheticobjectsegments,
title={SOS: Synthetic Object Segments Improve Detection, Segmentation, and Grounding},
author={Weikai Huang and Jieyu Zhang and Taoyang Jia and Chenhao Zheng and Ziqi Gao and Jae Sung Park and Ranjay Krishna},
year={2025},
eprint={2510.09110},
archivePrefix={arXiv}… See the full description on the dataset page: https://huggingface.co/datasets/weikaih/SOC-Training-Data-Visualization.Spatial-Visualization-Benchmark
Spatial Visualization Benchmark
This repository contains the Spatial Visualization Benchmark. The evaluation code is released on: wangst0181/Spatial-Visualization-Benchmark.
Dataset Description
The SpatialViz-Bench aims to evaluate the spatial visualization capabilities of multimodal large language models, which is a key component of spatial abilities. Targeting 4 sub-abilities of Spatial Visualization, including mental rotation, mental folding, visual penetration, and… See the full description on the dataset page: https://huggingface.co/datasets/PLM-Team/Spatial-Visualization-Benchmark.
