datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
modelnet40_normal_resampled-compresseds3dis-compressedpoi-benchmark
POI Benchmark: Multi-City Multimodal Points of Interest
A large-scale multimodal benchmark pairing Points of Interest (POIs) with street-view imagery, aerial grid photos, and satellite imagery across 10 major cities on 3 continents.
Cities
Beijing, Chengdu, Guangzhou, Hong Kong, Shanghai, Shenzhen, London, Melbourne, New York, Sydney.
Contents
Path
Size
Type
Description
metadata_aligned.tar
8.6 GB
11 JSON files
Enriched & aligned POI metadata per… See the full description on the dataset page: https://huggingface.co/datasets/sukiewang/poi-benchmark.nuscenes-compressedPointOdyssey_vlbm_old
PointOdyssey to VLBM Format Conversion Report
This document summarizes the process and results of converting the PointOdyssey training dataset to the Visual Lattice Boltzmann Model (VLBM) format.
Conversion Overview
The PointOdyssey train split was converted using a multi-processed Python script (pointodyssey2vlbm.py). The conversion involved transforming source RGB images, 16-bit depth maps, and coordinate annotations into the standardized format used by the VLBM dataset… See the full description on the dataset page: https://huggingface.co/datasets/ZhengGuangze/PointOdyssey_vlbm_old.pointsource_noisesThis repo was created to host only the pointsource noises, allowing for a download that does not include the RIRs, etc.
The original dataset containing all RIRs and noises can be downloaded by:
wget https://openslr.trmal.net/resources/28/rirs_noises.zip
Below is the relevant text from the original README.
This data includes all the pointsource noises
used in the paper "A Study on Data Augmentation of Reverberant Speech for Robust Speech Recognition"
submitted to ICASSP 2017.
Here are the… See the full description on the dataset page: https://huggingface.co/datasets/gfdb/pointsource_noises.
