datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mmu_ps1_sne_ia
mmu_ps1_sne_ia HATS Catalog Collection
This is the collection of HATS catalogs representing mmu_ps1_sne_ia.
This dataset is part of the Multimodal Universe,
a large-scale collection of multimodal astronomical data. For full details, see the paper:
The Multimodal Universe: Enabling Large-Scale Machine Learning with 100TBs of Astronomical Scientific Data.
Access the catalog
We recommend the use of the LSDB Python framework to access HATS catalogs.
LSDB can be… See the full description on the dataset page: https://huggingface.co/datasets/UniverseTBD/mmu_ps1_sne_ia.ps1_sne_ia---
description: 'Time-series dataset from the Pan-STARRS1 (PS1).
'
homepage: https://iopscience.iop.org/article/10.3847/1538-4357/aab9bb/pdf
version: 1.0.0
citation: "% % ACKNOWLEDGEMENTS\n% Here is the text for acknowledging PS1 in your \ publications:\n% \n% The Pan-STARRS1 Surveys (PS1) and the PS1 public science \ archive have been made possible through contributions by the Institute for Astronomy, \ the University of Hawaii, the Pan-STARRS Project Office, the Max-Planck Society \… See the full description on the dataset page: https://huggingface.co/datasets/MultimodalUniverse/ps1_sne_ia.ps1-discovery-corpus
PS1 Discovery Corpus
Canonical repo & full docs on GitHub →
A multilingual metadata corpus of 7,995 PlayStation 1 games, built to be searched
by feel — "a cozy fishing game with an anime aesthetic", "a 1996/97 Japanese game
where you could send letters", "a bleak sci-fi adventure nobody remembers" — rather than
by popularity or rigid filters. It is deliberately biased toward obscure and
Japan-exclusive titles: the long tail most databases skip.
The dataset's value is its… See the full description on the dataset page: https://huggingface.co/datasets/kenrinzero/ps1-discovery-corpus.ps1_sne_ia---
description: 'Time-series dataset from the Pan-STARRS1 (PS1).
'
homepage: https://iopscience.iop.org/article/10.3847/1538-4357/aab9bb/pdf
version: 1.0.0
citation: "% % ACKNOWLEDGEMENTS\n% Here is the text for acknowledging PS1 in your \ publications:\n% \n% The Pan-STARRS1 Surveys (PS1) and the PS1 public science \ archive have been made possible through contributions by the Institute for Astronomy, \ the University of Hawaii, the Pan-STARRS Project Office, the Max-Planck Society \… See the full description on the dataset page: https://huggingface.co/datasets/EiffL/ps1_sne_ia.h2f_ps1_pasquet_autolabeling_radius_norm_p6
Description
This is the spectroscopic galaxy sample of Pasquet et al. [1], originally built
for photometric redshift estimation from SDSS, re-imaged as multi-resolution
PanSTARRS cutouts through
hips2fits
(PanSTARRS DR1 HiPS, served by CDS).
It contains 639,750 examples drawn from 481,589 galaxies, built by
autolabeling: the cutouts are not centered on the galaxy but on a
simulated transient position. For each galaxy, positions are drawn
uniformly inside the elliptical footprint… See the full description on the dataset page: https://huggingface.co/datasets/PRISM-Astro/h2f_ps1_pasquet_autolabeling_radius_norm_p6.h2f_ps1_pasquet
Description
This is the spectroscopic galaxy sample of Pasquet et al. [1], originally built
for photometric redshift estimation from SDSS, re-imaged as multi-resolution
PanSTARRS cutouts through
hips2fits
(PanSTARRS DR1 HiPS, served by CDS).
It contains 481,589 galaxies, one row each. Every cutout is centered on the
galaxy itself, which makes this the direct counterpart of
h2f_ps1_pasquet_autolabeling_0_01pct:
same sample, same metadata and same folds, but there the cutouts are… See the full description on the dataset page: https://huggingface.co/datasets/PRISM-Astro/h2f_ps1_pasquet.h2f_ps1_pasquet_autolabeling_0_01pct
Description
This is the spectroscopic galaxy sample of Pasquet et al. [1], originally built
for photometric redshift estimation from SDSS, re-imaged as multi-resolution
PanSTARRS cutouts through
hips2fits
(PanSTARRS DR1 HiPS, served by CDS).
It contains 506.572 examples drawn from 481,589 galaxies, built by
autolabeling: the cutouts are not centered on the galaxy but on a
simulated transient position. For each galaxy, positions are drawn
uniformly inside the elliptical footprint… See the full description on the dataset page: https://huggingface.co/datasets/PRISM-Astro/h2f_ps1_pasquet_autolabeling_0_01pct.
