tao
Datasets
All datasets matching “tao”CRA5-Dataset
Climate science data can be compressed efficiently by dual-stage extreme compression with a variational auto-encoder transformer
Introduction and get started
CRA5 dataset now is available at OneDrive
Paper Summary
We introduce VAEformer, a variational autoencoder transformer designed for the extreme compression of climate data. Addressing the storage challenges of massive datasets like ERA5, VAEformer utilizes a… See the full description on the dataset page: https://huggingface.co/datasets/taohan10200/CRA5-Dataset.Dao_taoTaobao-ProductsAdsDBEmilia-Dataset-tokenisedTaobao-MMTAOBAO-MM: A Long Sequence Recommendation Dataset with Multimodal Embeddings at Scale
Overview |
Dataset Description |
Download and Use |
Contact |
Citation
Overview
TAOBAO-MM is a large-scale recommendation dataset derived from user interaction logs on Taobao, one of the world’s largest e-commerce platforms. The dataset features historical behavior sequences of up to 1,000 interactions per user and includes high-quality multimodal embeddings for… See the full description on the dataset page: https://huggingface.co/datasets/TaoBao-MM/Taobao-MM.LVOmniBench
LVOmniBench: Pioneering Long Audio-Video Understanding Evaluation for Omnimodal LLMs
LVOmniBench is a new audio-visual understanding evaluation benchmark in long-form audio-video inputs. 🌟
🔥 News
2026.03.19 🌟 We are very proud to launch LVOmniBench, the pioneering comprehensive evaluation benchmark of OmniLLMs in Long Audio-Video Understanding Evaluation!
✨ LVOmniBench Introduction
Recent advancements in omnimodal large language models… See the full description on the dataset page: https://huggingface.co/datasets/KD-TAO/LVOmniBench.
