datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
GeoMeld
🌍 GeoMeld Multi-Modal Earth Observation Dataset (WebDataset)
GeoMeld is a large-scale multi-modal remote sensing dataset introduced in our CVPRW 2026 paper on semantically grounded foundation modeling.
GeoMeld contains approximately 2.5 million spatially aligned samples spanning heterogeneous sensing modalities and spatial resolutions, paired with semantically grounded captions generated through an agentic pipeline.
The dataset is designed to support multimodal representation… See the full description on the dataset page: https://huggingface.co/datasets/vimageiitb/GeoMeld.ViMUL-Bench
ViMUL-Bench: A Culturally-diverse Multilingual Multimodal Video Benchmark
Overview
The evaluation toolkit to be used is lmms-eval. This toolkit facilitates the evaluation of models across multiple tasks and languages.
Key Features
🌍 14 Languages: English, Chinese, Spanish, French, German, Hindi, Arabic, Russian, Bengali, Urdu, Sinhala, Tamil, Swedish, Japanese🎭 15 Categories: Including 8 culturally diverse categories (lifestyles, festivals, foods… See the full description on the dataset page: https://huggingface.co/datasets/MBZUAI/ViMUL-Bench.vi-mmarcotrace-forge-kimi-k3-dry-v0
trace-forge-kimi-k3-dry-v0
Reasoning traces from kimi-k3 (Moonshot native API) over a
16-prompt self-authored bank, 2 samples per prompt,
generated on 2026-07-25. All numbers in this card are measured.
Author and maintainer: Vimal Nakrani (vimalnakrani), sole author and maintainer.
Configs
The raw config has all 32 records: prompt, final answer in
content, the model's reasoning in its own reasoning field,
finish_reason, verification status, token usage, seed, and… See the full description on the dataset page: https://huggingface.co/datasets/vimalnakrani/trace-forge-kimi-k3-dry-v0.context_instruct_vimhopfcvi_msa_sar
VN-SarMSA-vi
Vietnamese aspect-level sentiment (7-point, −3…+3) and sarcasm (binary) annotations
for hotel/restaurant reviews.
17,236 (sentence, aspect) pairs; splits 13,758 / 1,739 / 1,739 (train/validation/test)
Splits are leakage-free by construction: reviews sharing any normalized sentence are
grouped before splitting, so no sentence or review crosses splits
source column marks synthetic augmentation ('augmented'); evaluate sentiment on
source != 'augmented' for a… See the full description on the dataset page: https://huggingface.co/datasets/phamluan/vi_msa_sar.details_sometimesanotion__Qwen2.5-14B-Vimarckoso-v3
Dataset Card for Evaluation run of sometimesanotion/Qwen2.5-14B-Vimarckoso-v3
Dataset automatically created during the evaluation run of model sometimesanotion/Qwen2.5-14B-Vimarckoso-v3.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_sometimesanotion__Qwen2.5-14B-Vimarckoso-v3.gold_evidence_instruct_vimhopfcDatasetsPADvimedaqa-rft-poolvimedaqa-status-runsvimedaqa-bon-candidates
