datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hest
Model Card for HEST-1k
What is HEST-1k?
A collection of 1,276 spatial transcriptomic profiles, each linked and aligned to a Whole Slide Image (with pixel size < 1.15 µm/px) and metadata.
HEST-1k was assembled from 180 public and internal cohorts encompassing:
26 organs
2 species (Homo Sapiens and Mus Musculus)
398 cancer samples from 25 cancer types.
HEST-1k processing enabled the identification of >1.5 million expression/morphology pairs and >76 million nuclei… See the full description on the dataset page: https://huggingface.co/datasets/MahmoodLab/hest.hest-benchDeepSpot2Cell-HEST1k-Virtual-SingleCell
DeepSpot2Cell Virtual Single-Cell Spatial Transcriptomics
Virtual single-cell gene expression predictions for Visium spatial transcriptomics
samples, generated by DeepSpot2Cell.
Overview
This dataset provides predicted single-cell gene expression profiles for
Visium samples across 5,000 genes. The predictions were generated by running
a trained DeepSpot2Cell model on preprocessed Visium data from HEST-1k.
DeepSpot2Cell uses a permutation-invariant DeepSet… See the full description on the dataset page: https://huggingface.co/datasets/GravityBeng/DeepSpot2Cell-HEST1k-Virtual-SingleCell.aerial-isac-pusch-hest
Aerial ISAC PUSCH Channel Estimates
Uplink PUSCH DMRS channel estimates from a 5G CBRS cell running indoors on the
NVIDIA Aerial
testbed, paired with camera-derived floor positions of a person walking through the
cell: 1.19 million estimates over 18 runs, nine with a person in the area and nine
recorded empty as a background reference.
Plan view in the label coordinate frame, with the recorded track of a clear run
and of an obstacle run.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/aerial-isac-pusch-hest.hestenet-qa
Hestenet Question-Answer
The dataset is based on data from Hestenettet in the Danish Gigaword corpus.
Question-answer pairs are purely extracted on the basis of heuristics, and have not been manually evaluated.
The dataset was created for aiding the training of sentence transformer models in the Danish Foundation Models project.
The dataset is currently not production-ready.
More Information needed
hestenettet
Dataset Card for "hestenettet"
Subset of Gigaword.
https://huggingface.co/datasets/DDSC/partial-danish-gigaword-no-twitter
hest-1k-labeledhest-stimage-1k4mHEST_Visium_virtual_single_cell_transcriptomics
HEST-1k Visium virtual single-cell spatial transcriptomics
Predicted virtual single-cell gene expression for HEST-1k Visium samples,
produced with DeepSpot2Cell. DeepSpot2Cell is a Deep Sets model that predicts
transcriptomic profiles at single-cell resolution from H&E images and CellViT cell
segmentations, trained with spot-level Visium supervision only (no single-cell
ground truth). At inference every detected cell in each Visium spot receives its
own predicted expression… See the full description on the dataset page: https://huggingface.co/datasets/ratschlab/HEST_Visium_virtual_single_cell_transcriptomics.qsb-heston
QuantScenarioBench — Heston Benchmark Dataset
This dataset is a representative benchmark sample generated by QuantScenarioBench, a JAX-native framework for reproducible stochastic market scenario generation.
It contains 10,000 independent simulation paths under the Heston stochastic volatility model over 252 daily time steps (1 year horizon). Each path includes both the asset price trajectory (observation) and the instantaneous variance trajectory (latent state).
Need a larger… See the full description on the dataset page: https://huggingface.co/datasets/QuantScenarioBench/qsb-heston.HEST_Xenium_virtual_spatial_transcriptomics
HEST Xenium virtual spatial transcriptomics
This repository contains predicted spatial transcriptomics for HEST Xenium H&E
slides produced with DeepSpot-M.
Authors: Kalin Nonchev, Sebastian Dawo, Karina Silina, Viktor Hendrik
Koelzer, and Gunnar Rätsch.
Paper: DeepSpot-M: a multimodal foundation model for transcriptome-wide virtual spatial transcriptomics from histology (medRxiv, 2026; see the citation below).
Code: https://github.com/ratschlab/DeepSpotM.
News… See the full description on the dataset page: https://huggingface.co/datasets/ratschlab/HEST_Xenium_virtual_spatial_transcriptomics.HeStutters-labeled-3.6k
HeStutters - Annotated Stuttering Events for 3.6k Audio Clips
Adapted from the Apple SEP-28k dataset.
hestia_isitwrongtotrytopickupgirlsinadungeon
Dataset of hestia (Dungeon ni Deai wo Motomeru no wa Machigatteiru no Darou ka)
This is the dataset of hestia (Dungeon ni Deai wo Motomeru no wa Machigatteiru no Darou ka), containing 200 images and their tags.
Images are crawled from many sites (e.g. danbooru, pixiv, zerochan ...), the auto-crawling system is powered by DeepGHS Team(huggingface organization).
laurashin_SECs_Hester_Peirce_Tackles_Frustrating_Crypto_Regulation__And_Why_Its_So_Slow__-_Ep__5hest-websitehest1k-organ-subsethest_bench_trainhesti-v1-datasethest_bench_testhest_st_filestwitter-HesterJack22271-2026.01.04-2007783137116594578-LdJNDkHXhLJ9P3tU-part1hestphoenix_hestdata
