datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hestenettet
Dataset Card for "hestenettet"
Subset of Gigaword.
https://huggingface.co/datasets/DDSC/partial-danish-gigaword-no-twitter
hestenet-qa
Hestenet Question-Answer
The dataset is based on data from Hestenettet in the Danish Gigaword corpus.
Question-answer pairs are purely extracted on the basis of heuristics, and have not been manually evaluated.
The dataset was created for aiding the training of sentence transformer models in the Danish Foundation Models project.
The dataset is currently not production-ready.
More Information needed
HeStutters-labeled-3.6k
HeStutters - Annotated Stuttering Events for 3.6k Audio Clips
Adapted from the Apple SEP-28k dataset.
qsb-heston
QuantScenarioBench — Heston Benchmark Dataset
This dataset is a representative benchmark sample generated by QuantScenarioBench, a JAX-native framework for reproducible stochastic market scenario generation.
It contains 10,000 independent simulation paths under the Heston stochastic volatility model over 252 daily time steps (1 year horizon). Each path includes both the asset price trajectory (observation) and the instantaneous variance trajectory (latent state).
Need a larger… See the full description on the dataset page: https://huggingface.co/datasets/QuantScenarioBench/qsb-heston.laurashin_SECs_Hester_Peirce_Tackles_Frustrating_Crypto_Regulation__And_Why_Its_So_Slow__-_Ep__5hest_bench_trainhest_bench_test
