CoolFace
18 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01LevinHarness /rfantibody-assets LevinHarness/rfantibody-assets — public mirror of third-party runtime assets This dataset is a public mirror of third-party runtime assets required by the Levin Harness plugin(s) listed below, mirrored verbatim from their original sources with SHA-256 pinning. It is not an official distribution: nothing here is published under this account's own terms, and it is not affiliated with or endorsed by any upstream project. Ownership and licensing Every file remains… See the full description on the dataset page: https://huggingface.co/datasets/LevinHarness/rfantibody-assets.textn<1K0 likes167 downloads11d agoHugging Face02rfab85 /crypto-5s-market-data-adausdc-sample ADA/USDC High-Frequency Market Microstructure Data Free 7-Day Sample This repository provides a free 7-day sample of a much larger privately collected high-frequency cryptocurrency market dataset. The complete historical archive contains millions of market snapshots, with data collection starting in December 2025, across 12 crypto/USDC markets. The public ADA/USDC sample contains: 81,579 market snapshots 97 columns 7 days of continuous historical data 20 bid + 20… See the full description on the dataset page: https://huggingface.co/datasets/rfab85/crypto-5s-market-data-adausdc-sample.tabular10K<n<100K1 likes82 downloads1mo agoHugging Face03multimolecule /rfam Rfam Rfam is a database of structure-annotated multiple sequence alignments, covariance models and family annotation for a number of non-coding RNA, cis-regulatory and self-splicing intron families. The seed alignments are hand curated and aligned using available sequence and structure data, and covariance models are built from these alignments using the INFERNAL v1.1.4 software suite. The full regions list is created by searching the RFAMSEQ database using the covariance model… See the full description on the dataset page: https://huggingface.co/datasets/multimolecule/rfam.texttext-generation10M<n<100M2 likes68 downloads1y agoHugging Face04sgetttt /rfantibody-assets LevinHarness/rfantibody-assets — public mirror of third-party runtime assets This dataset is a public mirror of third-party runtime assets required by the Levin Harness plugin(s) listed below, mirrored verbatim from their original sources with SHA-256 pinning. It is not an official distribution: nothing here is published under this account's own terms, and it is not affiliated with or endorsed by any upstream project. Ownership and licensing Every file remains… See the full description on the dataset page: https://huggingface.co/datasets/sgetttt/rfantibody-assets.textn<1K0 likes51 downloads11d agoHugging Face05freococo /rfa_rakhine_language_voices RFA Rakhine Language Voices This dataset contains 14.53 hours of audio in the Rakhine (Arakanese) language, sourced from news broadcasts by Radio Free Asia (RFA) Burmese. This is one of the largest publicly accessible audio resources for the Rakhine language, designed to support research in low-resource automatic speech recognition (ASR), voice activity detection, and other speech-related tasks. The audio has been automatically segmented into manageable chunks and prepared in the… See the full description on the dataset page: https://huggingface.co/datasets/freococo/rfa_rakhine_language_voices.audioautomatic-speech-recognition1K<n<10K0 likes41 downloads1y agoHugging Face06afg1 /rfam-taxonomy-lookuptext10M<n<100M0 likes41 downloads11d agoHugging Face07freococo /rfa_shan_language_voices RFA Shan Language Voices This dataset contains 20.58 hours of audio in the Shan (Tai-Yai) language, sourced from news broadcasts by Radio Free Asia (RFA) Burmese. This is one of the largest publicly accessible audio resources for the Shan language, designed to support research in low-resource automatic speech recognition (ASR), voice activity detection, and other speech-related tasks. The audio has been automatically segmented into 5,047 manageable chunks and prepared in the… See the full description on the dataset page: https://huggingface.co/datasets/freococo/rfa_shan_language_voices.audioautomatic-speech-recognition1K<n<10K0 likes30 downloads1y agoHugging Face08rFathi03 /Arabic_Diacritized_Audio_Datasetaudio1K<n<10K1 likes21 downloads2y agoHugging Face09rFathi03 /Diacritized-Case-Ending-Errors-TTT-V2.0textn<1K0 likes12 downloads1y agoHugging Face10rFathi03 /Diacritized-Case-Ending-Errors-TTT-V3.0text1K<n<10K0 likes12 downloads1y agoHugging Face11rFathi03 /Diacritized-Case-Ending-Errors-TTT-V1.0textn<1K0 likes11 downloads1y agoHugging Face12rfazal853ml /Meal_Planner_2kplusdataset Meal Planner 2000+ sampkle dataset text1K<n<10K0 likes9 downloads6mo agoHugging Face13Sofoklis /rfam_sample_padded_arraystext1K<n<10K0 likes7 downloads2y agoHugging Face14rFathi03 /Models-Comparison-Datasettextn<1K0 likes3 downloads1y agoHugging Face15rFathi03 /Clean-vs-Undiacritized-Datasettextn<1K0 likes3 downloads1y agoHugging Face16rfazal853ml /Meal_plan_dataset Meal planing dataset text1K<n<10K0 likes3 downloads6mo agoHugging Face17rfazal853ml /meal_plan_dataset_expert meal_plan_dataset_expert textn<1K0 likes3 downloads5mo agoHugging Face18rfazal853ml /FoodieFlow Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/rfazal853ml/FoodieFlow.textn<1K0 likes2 downloads6mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.