CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01HANZO2345 /PS1HHHHHHHHHLDFGLDFLGDFLGLDFGLDFGLDLVFDVLDFLGDFG0 likes287 downloads13d agoHugging Face02UniverseTBD /mmu_ps1_sne_ia mmu_ps1_sne_ia HATS Catalog Collection This is the collection of HATS catalogs representing mmu_ps1_sne_ia. This dataset is part of the Multimodal Universe, a large-scale collection of multimodal astronomical data. For full details, see the paper: The Multimodal Universe: Enabling Large-Scale Machine Learning with 100TBs of Astronomical Scientific Data. Access the catalog We recommend the use of the LSDB Python framework to access HATS catalogs. LSDB can be… See the full description on the dataset page: https://huggingface.co/datasets/UniverseTBD/mmu_ps1_sne_ia.tabularn<1K0 likes157 downloads4mo agoHugging Face03HTayy /LinearSpectre_PS1_Cifar100_alpha_0portiontextn<1K0 likes99 downloads4mo agoHugging Face04jpq-repro /msmarco-passage-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__ msmarco-passage-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__ Description This is the PyTerrier JPQIndex for MSMARCO v1 passage corpus, which corresponds to an result from the SIGIR 2026 reproducibility paper. Usage # Load the artifact import pyterrier as pt import pyterrier_dr index = pt.Artifact.from_hf('jpq-repro/msmarco-passage-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__') model =… See the full description on the dataset page: https://huggingface.co/datasets/jpq-repro/msmarco-passage-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__.text-retrieval0 likes61 downloads2mo agoHugging Face05jpq-repro /nq-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__ nq-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__ Description This is the PyTerrier JPQIndex for the Wikipedia 2018 corpus used by Natural Questions (NQ). Usage # Load the artifact import pyterrier as pt import pyterrier_dr index = pt.Artifact.from_hf('jpq-repro/nq-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__') model = pyterrier_dr.TctColBert.hnp() model.model =… See the full description on the dataset page: https://huggingface.co/datasets/jpq-repro/nq-train__tct_colbert__faiss2opq__M96_nbits8__ps159744__neg200__ibn__lr__.text-retrieval1 likes61 downloads2mo agoHugging Face06HTayy /LinearSpectre_PS1_Cifar100_only_gla_4textn<1K0 likes54 downloads4mo agoHugging Face07HTayy /LinearSpectre_PS1_Cifar100_not_applytextn<1K1 likes48 downloads5mo agoHugging Face08Sandy-sys /aditi-ps18-media ADITI PS-18 — source references (not a data mirror) Datasets are referenced, never redistributed. This repo carries a pinned manifest — repo id + exact commit SHA + file patterns + licence — so the corpus is reproducible byte-for-byte without re-hosting anyone's data. Why: several sources carry platform terms that forbid redistribution (Reddit/ConvoKit, Twitter-derived, ShareChat/Moj). The acquisition policy for this programme is explicit — "do not redistribute raw social-media… See the full description on the dataset page: https://huggingface.co/datasets/Sandy-sys/aditi-ps18-media.0 likes45 downloads1mo agoHugging Face09HTayy /LinearSpectre_PS1_Cifar100_alpha_100portiontextn<1K0 likes43 downloads5mo agoHugging Face10HTayy /LinearSpectre_PS1_Cifar100_all_blockstextn<1K0 likes38 downloads5mo agoHugging Face11HTayy /LinearSpectre_PS1_Cifar10_only_glatextn<1K0 likes35 downloads5mo agoHugging Face12HTayy /LinearSpectre_PS1_Cifar100_only_glatextn<1K0 likes30 downloads5mo agoHugging Face13HTayy /LinearSpectre_PS1_Cifar10_only_gla_2textn<1K0 likes28 downloads5mo agoHugging Face14HTayy /LinearSpectre_PS1_Cifar10_only_gla_3textn<1K0 likes21 downloads4mo agoHugging Face15HTayy /LinearSpectre_PS1_Cifar100_only_gla_3textn<1K0 likes20 downloads5mo agoHugging Face16HTayy /LinearSpectre_PS1_Cifar10_alpha_00portiontextn<1K0 likes19 downloads4mo agoHugging Face17HTayy /Performers_PS1_Cifar100_warmuptextn<1K0 likes18 downloads7mo agoHugging Face18MultimodalUniverse /ps1_sne_ia--- description: 'Time-series dataset from the Pan-STARRS1 (PS1). ' homepage: https://iopscience.iop.org/article/10.3847/1538-4357/aab9bb/pdf version: 1.0.0 citation: "% % ACKNOWLEDGEMENTS\n% Here is the text for acknowledging PS1 in your \ publications:\n% \n% The Pan-STARRS1 Surveys (PS1) and the PS1 public science \ archive have been made possible through contributions by the Institute for Astronomy, \ the University of Hawaii, the Pan-STARRS Project Office, the Max-Planck Society \… See the full description on the dataset page: https://huggingface.co/datasets/MultimodalUniverse/ps1_sne_ia.tabularn<1K0 likes17 downloads2y agoHugging Face19HTayy /LinearSpectre_PS1_Cifar10_alpha_100portiontextn<1K0 likes16 downloads5mo agoHugging Face20Hyperccino /Any-to-PS1-v1.0This is a dataset consisting of 17 image pairs for finetuning to convert an image subject -> retro PS1 style. imageimage-to-imagen<1K0 likes16 downloads5mo agoHugging Face21HTayy /LinearSpectre_PS1_Cifar10_not_apply_2textn<1K0 likes15 downloads4mo agoHugging Face22kenrinzero /ps1-discovery-corpus PS1 Discovery Corpus &nbsp; Canonical repo & full docs on GitHub → A multilingual metadata corpus of 7,995 PlayStation 1 games, built to be searched by feel — "a cozy fishing game with an anime aesthetic", "a 1996/97 Japanese game where you could send letters", "a bleak sci-fi adventure nobody remembers" — rather than by popularity or rigid filters. It is deliberately biased toward obscure and Japan-exclusive titles: the long tail most databases skip. The dataset's value is its… See the full description on the dataset page: https://huggingface.co/datasets/kenrinzero/ps1-discovery-corpus.tabulartext-retrieval1K<n<10K0 likes15 downloads3mo agoHugging Face23victkk /ps12k Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/victkk/ps12k.text10K<n<100K0 likes13 downloads1y agoHugging Face24HTayy /LinearSpectre_PS1_Cifar10_all_blockstextn<1K1 likes12 downloads5mo agoHugging Face25HTayy /LinearSpectre_PS1_Cifar100_only_ffttextn<1K0 likes12 downloads5mo agoHugging Face26HTayy /LinearSpectre_PS1_Cifar10_not_applytextn<1K0 likes11 downloads5mo agoHugging Face27Hyperccino /PS1-to-Untextured-v1.0This is a dataset consisting of 16 image pairs of PS1 style models -> untextured model. This dataset is different than my Any-to-PS1-Untextured-v1.0 dataset in that this is specifically made for PS1 style inputs (specifically in the output style of my Qwen-Edit-2511-PS1-v1.0 model), and that this has more optimized image sizes. imageimage-to-imagen<1K0 likes11 downloads5mo agoHugging Face28PS-123 /LungNodule Dataset Card for "LungNodule" More Information needed image1K<n<10K0 likes9 downloads3y agoHugging Face29ps1293 /job_descriptiontextn<1K2 likes8 downloads3y agoHugging Face30EiffL /ps1_sne_ia--- description: 'Time-series dataset from the Pan-STARRS1 (PS1). ' homepage: https://iopscience.iop.org/article/10.3847/1538-4357/aab9bb/pdf version: 1.0.0 citation: "% % ACKNOWLEDGEMENTS\n% Here is the text for acknowledging PS1 in your \ publications:\n% \n% The Pan-STARRS1 Surveys (PS1) and the PS1 public science \ archive have been made possible through contributions by the Institute for Astronomy, \ the University of Hawaii, the Pan-STARRS Project Office, the Max-Planck Society \… See the full description on the dataset page: https://huggingface.co/datasets/EiffL/ps1_sne_ia.tabularn<1K0 likes8 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.