datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
marine-animals-multimodal-dataset
Marine Animals Multimodal Dataset 🐋
A comprehensive multimodal dataset combining audio recordings and images of 32 marine species.
Dataset Summary
Total samples: 24,911
Species: 32
Audio files: 1,357 unique recordings
Images: 581 (309 matched + 272 from iNaturalist)
Features
species (string): Species name
label (int32): Numeric label (0–31)
audio (Audio): Audio recording of the species
image (Image): Species image
image_index (int32): Image number… See the full description on the dataset page: https://huggingface.co/datasets/Hariprasath5128/marine-animals-multimodal-dataset.Open-o3-Video
Open-o3 Video
TL; DR: Open-o3 Video integrates explicit spatio-temporal evidence into video reasoning through curated STGR datasets and a two-stage SFT–RL training strategy, achieving state-of-the-art results on V-STAR and delivering verifiable, reliable reasoning for video understanding.
Data
To provide unified spatio-temporal supervision for grounded video reasoning, we build two datasets: STGR-CoT-30k for supervised fine-tuning and STGR-RL-36k for reinforcement… See the full description on the dataset page: https://huggingface.co/datasets/marinero4972/Open-o3-Video.marine-figuresMarineLife-16K
Dataset Card for MarineLife-16K
We introduce MarineLife-16K, a marine-domain video benchmark designed to evaluate the video understanding capabilities of Vision-Language Models (VLMs). MarineLife-16K contains 2,000 video-text pairs and 16,080 video-question-answer pairs across a collection of 2,000 marine videos, including 12,080 multiple-choice questions and 4,000 open-ended questions. The benchmark emphasizes specialized marine knowledge, visual reasoning, temporal… See the full description on the dataset page: https://huggingface.co/datasets/MarineLife-16K/MarineLife-16K.marine-ecological-feature-dataset
Marine Ecological Feature Dataset Assets
"
"This dataset repository contains currently available assets for the typical marine ecological environment "
"feature recognition project.
"
"Current uploaded assets are partial and include GF1/GF2 scene inputs available locally plus the GF6 full-scene "
"inference mask/preview generated by the current prototype model. The normalized manifest workflow is defined "
"in dataset_standard.md; a… See the full description on the dataset page: https://huggingface.co/datasets/cuibinge/marine-ecological-feature-dataset.MARINER
MARINER
A maritime ship object detection dataset with 63 fine-grained ship categories.
Dataset Structure
test/
├── test.json # Bounding box annotations
├── *.jpg # Images (1000 total)
Annotation Format
Each entry in test.json:
{
"image_name": "054A_109.jpg",
"grounding_info": {
"text_prompt": "ship .",
"objects": [
{
"label": "054A",
"score": 0.9291,
"box_norm": [x_min, y_min, x_max… See the full description on the dataset page: https://huggingface.co/datasets/moinn1/MARINER.LLM-Vision-Marine-Animals
Dataset Card for Benchmarking Large Language Models for Image Classification of Marine Mammals
As Artificial Intelligence (AI) has developed rapidly over the past few decades, the new generation of AI, Large Language Models (LLMs) trained on massive datasets, has achieved ground-breaking performance in many applications. Further progress has been made in multimodal LLMs, with many datasets created to evaluate LLMs with vision abilities. However, none of those datasets focuses… See the full description on the dataset page: https://huggingface.co/datasets/yeyimilk/LLM-Vision-Marine-Animals.MARINER
MARINER
A maritime ship object detection dataset with 63 fine-grained ship categories.
Dataset Structure
test/
├── test.json # Bounding box annotations
├── *.jpg # Images (1000 total)
Annotation Format
Each entry in test.json:
{
"image_name": "054A_109.jpg",
"grounding_info": {
"text_prompt": "ship .",
"objects": [
{
"label": "054A",
"score": 0.9291,
"box_norm": [x_min, y_min, x_max, y_max]… See the full description on the dataset page: https://huggingface.co/datasets/viviwang/MARINER.marine-animals-multimodalmarine_vla_dataset
Marine VLA Dataset
Vision-Language-Action dataset for autonomous marine vessel navigation using SmolVLA.
Dataset Structure
LeRobot-style format with 37 episodes, 12175 frames:
data/
episode_000000/
episode_data.json # frame-by-frame labels + metadata
observation.images.camera_0/
000000.jpg # 640x480 RGB frames
000001.jpg
...
episode_000001/
...
dataset_info.json # schema, label names, stats… See the full description on the dataset page: https://huggingface.co/datasets/MSaalaamaa/marine_vla_dataset.TW_Marine_2cls_datasetts-aims-reefscapes-marine-featuresmarine-species-zh
海纳海洋生物数据集 | Marine Species Dataset (Chinese)
4662 个海洋物种的结构化数据集:学名、中文俗名、完整分类阶元、命名人、OBIS 分布记录数、图片链接(逐行标注授权协议与作者署名)、中文简介。
A structured dataset of 4,662 marine species: scientific names, Chinese vernacular names, full taxonomy, authorities, OBIS occurrence counts, image links (with per-row license and author attribution), and 2,120 Chinese descriptions translated/organized from Chinese Wikipedia.
数据来自海洋科普公益平台 海纳 · hainahub.cn 的物种图鉴底层,随图鉴扩容滚动更新。
数据规模 / Stats
指标… See the full description on the dataset page: https://huggingface.co/datasets/hainahub/marine-species-zh.TW_Marine_5cls_datasetMarineMISR
MarineMISR
MarineMISR is a multi-image super-resolution (MISR) dataset pairing stacks of
Landsat 8/9 scenes (low-resolution, 30 m) with a single co-located Sentinel-2
scene (high-resolution, 10 m) over coastal and marine habitats. Each sample is
a 512x512 pixel patch (5.12 km x 5.12 km) sampled from one of three habitat
types — coral reef, seagrass, and mangrove — so the dataset can be
used to train and evaluate super-resolution models specifically over these
ecologically… See the full description on the dataset page: https://huggingface.co/datasets/Ysobel/MarineMISR.Aussys_Marine_DatasetTW_Marine_5cls_dataset-backupMarineOilSpillMarine-Semantic-Segmentation-Training-Datasetmarine_science_identifyNovelAI_marine_dataset
