datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
CanadaFireSat
Dataset Card for CanadaFireSat 🔥🛰️
In this benchmark, we investigate the potential of deep learning with multiple modalities for high-resolution wildfire forecasting. Leveraging different data settings across two types of model architectures: CNN-based and ViT-based.
📝 Published paper from ISPRS (ArXiv Version)
💿 Dataset repository on GitHub
🤖 Model repository on GitHub & Weights on Hugging Face
🟰 Another "Raw" version of the data with NPY files organized in different… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/CanadaFireSat.coralscapes
Coralscapes Dataset
The Coralscapes dataset is the first general-purpose dense semantic segmentation dataset for coral reefs. Similar in scope and with the same structure as the widely used Cityscapes dataset for urban scene understanding, Coralscapes allows for the benchmarking of semantic segmentation models in a new challenging domain.
Dataset Structure
The Coralscapes dataset spans 2075 images at 1024×2048px resolution… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/coralscapes.SwissView
Dataset Card for SwissView Dataset
Project Page
https://limirs.github.io/GeoExplorer/
GeoExplorer: Active Geo-localization with Curiosity-Driven Exploration
Dataset Summary
This dataset consists of two subsets:
SwissViewMonuments: which includes 15 images of atypical or distinctive scenes, such as unusual buildings and landscapes, with corresponding ground level images.
SwissView100, which comprises 100 images randomly selected from across the Swiss territory… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/SwissView.ValaisCD
ValaisCD Dataset
High-Resolution Aerial Change Detection (Switzerland, 2017–2023)
Project page: https://manonbechaz.github.io/2Player/
🗺️ Overview
ValaisCD is a high-resolution change detection dataset built from SwissTopo SWISSIMAGE 10 cm aerial imagery, covering several urban and peri-urban regions of the canton of Valais, Switzerland.It provides pairs of aerial images captured in 2017 and 2023, along with automatically generated building-change labels derived from… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/ValaisCD.CanadaFireSat-Raw
Dataset Card for CanadaFireSat 🔥🛰️
In this benchmark, we investigate the potential of deep learning with multiple modalities for high-resolution wildfire forecasting. Leveraging different data settings across two types of model architectures: CNN-based and ViT-based.
📝 Published paper from ISPRS (ArXiv Version)
💿 Dataset repository on GitHub
🤖 Model repository on GitHub & Weights on Hugging Face
🟰 Another "Clean" version of the data with PARQUET files can be found at… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/CanadaFireSat-Raw.EcoWikiRS
EcoWikiRS: Learning Ecological Representations of Satellite Images from Weak Supervision with Species Observations and Wikipedia
AuthorsValerie Zermatten · Javiera Castillo-Navarro · Pallavi Jain · Devis Tuia · Diego Marcos
Overview
The WikiRS dataset, composed of triplets of images, species list and Wikipedia sentences :
91k high-resolution aerial images (50cm, RGB bands) from the swissIMAGE product
crowd-sourced species observations from 2745 different… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/EcoWikiRS.details_paloalma__ECE-TW3-JRGL-V1HRSCD_clean
📚 HRSCD-Clean Dataset
Project page: https://manonbechaz.github.io/2Player/
📝 Description
HRSCD-Clean is a refined and higher-quality version of the original HRSCD remote-sensingchange detection dataset (Daudt et al., 2019). The dataset contains 291 bi-temporal aerialimage pairs, each at 10,000 × 10,000 px and 0.5 m spatial resolution, covering theregions of Rennes and Caen, France. Each pair is accompanied by a binary change mask and segmentation masks for both images.… See the full description on the dataset page: https://huggingface.co/datasets/EPFL-ECEO/HRSCD_clean.SMBU-ECE-GRADE1-MATERIALece-6514-SFT-LLM-A1-advancedbrgx53__3Blarenegv2-ECE-PRYMMAL-Martial-details
Dataset Card for Evaluation run of brgx53/3Blarenegv2-ECE-PRYMMAL-Martial
Dataset automatically created during the evaluation run of model brgx53/3Blarenegv2-ECE-PRYMMAL-Martial
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/brgx53__3Blarenegv2-ECE-PRYMMAL-Martial-details.lesubra__ECE-EIFFEL-3B-details
Dataset Card for Evaluation run of lesubra/ECE-EIFFEL-3B
Dataset automatically created during the evaluation run of model lesubra/ECE-EIFFEL-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/lesubra__ECE-EIFFEL-3B-details.EC_ESMFold
EC Dataset with ESMFold Structural Sequence
Description: The Enzyme Commission number (EC number) is a numerical classification scheme for enzymes, based on the chemical reactions they catalyze.
Number of labels: 585
Problem Type: multi_label_classification
Columns:
aa_seq: protein amino acid sequence
foldseek_seq: foldseek 20 3di structural sequence
ss8_seq: DSSP 8 secondary structure sequence
Github
Simple, Efficient and Scalable Structure-aware Adapter Boosts… See the full description on the dataset page: https://huggingface.co/datasets/AI4Protein/EC_ESMFold.LilRg__ECE_Finetunning-details
Dataset Card for Evaluation run of LilRg/ECE_Finetunning
Dataset automatically created during the evaluation run of model LilRg/ECE_Finetunning
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__ECE_Finetunning-details.lesubra__ECE-PRYMMAL-3B-SLERP-V1-details
Dataset Card for Evaluation run of lesubra/ECE-PRYMMAL-3B-SLERP-V1
Dataset automatically created during the evaluation run of model lesubra/ECE-PRYMMAL-3B-SLERP-V1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/lesubra__ECE-PRYMMAL-3B-SLERP-V1-details.Youlln__ECE-PRYMMAL0.5B-Youri-details
Dataset Card for Evaluation run of Youlln/ECE-PRYMMAL0.5B-Youri
Dataset automatically created during the evaluation run of model Youlln/ECE-PRYMMAL0.5B-Youri
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Youlln__ECE-PRYMMAL0.5B-Youri-details.Lil-R__PRYMMAL-ECE-1B-SLERP-V1-details
Dataset Card for Evaluation run of Lil-R/PRYMMAL-ECE-1B-SLERP-V1
Dataset automatically created during the evaluation run of model Lil-R/PRYMMAL-ECE-1B-SLERP-V1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__PRYMMAL-ECE-1B-SLERP-V1-details.Lil-R__2_PRYMMAL-ECE-7B-SLERP-details
Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP
Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-details.Marsouuu__lareneg3Bv2-ECE-PRYMMAL-Martial-detailsMarsouuu__general3Bv2-ECE-PRYMMAL-Martial-details
Dataset Card for Evaluation run of Marsouuu/general3Bv2-ECE-PRYMMAL-Martial
Dataset automatically created during the evaluation run of model Marsouuu/general3Bv2-ECE-PRYMMAL-Martial
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Marsouuu__general3Bv2-ECE-PRYMMAL-Martial-details.eCeLLM-Kor-gemma-templateLilRg__PRYMMAL-ECE-7B-SLERP-V3-details
Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V3
Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V3-details.LilRg__PRYMMAL-ECE-7B-SLERP-V6-details
Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V6
Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V6
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V6-details.LilRg__PRYMMAL-ECE-7B-SLERP-V7-details
Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V7
Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V7
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V7-details.Lil-R__2_PRYMMAL-ECE-7B-SLERP-V1-details
Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP-V1
Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP-V1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-V1-details.LilRg__PRYMMAL-ECE-7B-SLERP-V5-details
Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V5
Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V5
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V5-details.Lil-R__2_PRYMMAL-ECE-7B-SLERP-V2-details
Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP-V2
Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP-V2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-V2-details.Lil-R__PRYMMAL-ECE-7B-SLERP-V8-details
Dataset Card for Evaluation run of Lil-R/PRYMMAL-ECE-7B-SLERP-V8
Dataset automatically created during the evaluation run of model Lil-R/PRYMMAL-ECE-7B-SLERP-V8
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__PRYMMAL-ECE-7B-SLERP-V8-details.LilRg__PRYMMAL-ECE-7B-SLERP-V4-details
Dataset Card for Evaluation run of LilRg/PRYMMAL-ECE-7B-SLERP-V4
Dataset automatically created during the evaluation run of model LilRg/PRYMMAL-ECE-7B-SLERP-V4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/LilRg__PRYMMAL-ECE-7B-SLERP-V4-details.Lil-R__2_PRYMMAL-ECE-7B-SLERP-V3-details
Dataset Card for Evaluation run of Lil-R/2_PRYMMAL-ECE-7B-SLERP-V3
Dataset automatically created during the evaluation run of model Lil-R/2_PRYMMAL-ECE-7B-SLERP-V3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Lil-R__2_PRYMMAL-ECE-7B-SLERP-V3-details.
