OCEAN
Datasets
All datasets matching “OCEAN”OceanDepths
OceanDepths GeoTIFF Raster and Aligned ARGO Dataset
This dataset package contains the model-ready Ocean variables (ARGO submarine data, sea surface height, sea surface temperature and
salinity, as well as GLORYS reanalysis information for 50 depth levels. The ARGO data has been projected onto the GLORYS grid in order
to build a ML-ready dataset. The intention is that users can create tensors easily for CV-inspired ML approaches to ocean-variable
reconstruction. While… See the full description on the dataset page: https://huggingface.co/datasets/ESA-philab/OceanDepths.OceanCorpus
OceanCorpus
Dataset Description
OceanCorpus is a large-scale, multimodal dataset designed to inject structured marine domain knowledge into Large Language Models (LLMs). It aggregates data from three primary sources to support text generation, instruction tuning, and vision-language alignment:
Web Knowledge (Text-Only): A dataset of 113,626 instruction-style QA pairs extracted from Wikipedia and authoritative marine websites, available in Web/data.csv.
Paper… See the full description on the dataset page: https://huggingface.co/datasets/zjunlp/OceanCorpus.OceanTACO
Dataset Card: OceanTACO
Dataset Summary
This dataset is a multi-source collection of global ocean sea surface measurements, integrating numerical model reanalysis, L4 gap-filled products, L3 satellite observations, and in-situ data. The collection includes sea surface height (SSH), temperature (SST), salinity (SSS), wind speed, and other variables.
The L3 SWOT data has been processed onto a consistent regular grid through irreversible interpolation and coordinate… See the full description on the dataset page: https://huggingface.co/datasets/nilsleh/OceanTACO.oceanOcean
OCEAN Big Five Personality Dataset
Dataset Summary
OCEAN Big Five Personality Dataset contains structured text samples and questionnaires annotated with OCEAN (Openness, Conscientiousness, Extraversion, Agreeableness, Neuroticism) Big Five personality trait scores.
Dataset Structure
Language: English (en)
Fields: Text prompts, personality trait ratings.
planktonzilla-17M
Planktonzilla-17M Dataset
Overview
planktonzilla-17M is a large-scale, comprehensive dataset combining 17 million plankton images from all publicly available -to the best
of our knowledge- labeled plankton datasets. This unified collection enables researchers to train robust deep learning models for plankton
identification and classification across diverse imaging systems and oceanographic environments.
Each image includes a standardized taxonomic hierarchy… See the full description on the dataset page: https://huggingface.co/datasets/project-oceania/planktonzilla-17M.
