CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01raei /Schalk-2009-EEGMotorMovementImagery EEG Motor Movement/Imagery Dataset This is an unofficial mirror of the PhysioNet EEG Motor Movement/Imagery Dataset, version 1.0.0. It is not affiliated with or endorsed by the dataset contributors, their institutions, or PhysioNet. Source and documentation Original dataset: PhysioNet EEGMMIDB v1.0.0 Source version: 1.0.0, published September 9, 2009 Original publication: Schalk et al., IEEE Transactions on Biomedical Engineering (2004) BCI2000 project:… See the full description on the dataset page: https://huggingface.co/datasets/raei/Schalk-2009-EEGMotorMovementImagery.documentn<1K0 likes13k downloads22d agoHugging Face02HAERAE-HUB /HAE_RAE_BENCH_1.1The HAE_RAE_BENCH 1.1 is an ongoing project to develop a suite of evaluation tasks designed to test the understanding of models regarding Korean cultural and contextual nuances. Currently, it comprises 13 distinct tasks, with a total of 4900 instances. Please note that although this repository contains datasets from the original HAE-RAE BENCH paper, the contents are not completely identical. Specifically, the reading comprehension subset from the original version has been removed due to… See the full description on the dataset page: https://huggingface.co/datasets/HAERAE-HUB/HAE_RAE_BENCH_1.1.textmultiple-choice1K<n<10K20 likes3.5k downloads2y agoHugging Face03raei /Jeong-2022-InternationalBCICompetition2020Review-track3 BCI Competition 2020 Track 3: imagined speech classification This is an unofficial mirror of Track 3 only from the 2020 International BCI Competition, described by Jeong et al. (2022). It is not affiliated with or endorsed by the authors, their institutions, or OSF. Source and attribution Original data: 2020 International BCI Competition, OSF project pq7vb, folder Track#3 Imagined speech classification. Paper: Jeong et al., 2020 International brain–computer… See the full description on the dataset page: https://huggingface.co/datasets/raei/Jeong-2022-InternationalBCICompetition2020Review-track3.documentn<1K0 likes1.1k downloads19d agoHugging Face04raeidsaqur /NIFTY The News-Informed Financial Trend Yield (NIFTY) Dataset. The News-Informed Financial Trend Yield (NIFTY) Dataset. Details of the dataset, including data procurement and filtering can be found in the paper here: https://arxiv.org/abs/2405.09747. For the NIFTY-RL LLM alignment dataset please use nifty-rl. 📋 Table of Contents 🧩 NIFTY Dataset 📋 Table of Contents 📖 Usage Downloading the dataset Dataset structure Large Language Models ✍️ Contributing 📝 Citing 🙏… See the full description on the dataset page: https://huggingface.co/datasets/raeidsaqur/NIFTY.textmultiple-choice1K<n<10K14 likes309 downloads2y agoHugging Face05HAERAE-HUB /HAE_RAE_BENCH_1.0The HAE_RAE_BENCH 1.0 is the original implementation of the dataset froom the paper: HAE-RAE BENCH paper. The benchmark is a collection of 1,538 instances across 6 tasks: standard_nomenclature, loan_word, rare_word, general_knowledge, history and reading comprehension. To replicate the studies from the paper, see below. Dataset Overview Task Instances Version Explanation standard_nomenclature 153 v1.0 Multiple-choice questions about Korean standard nomenclatures from… See the full description on the dataset page: https://huggingface.co/datasets/HAERAE-HUB/HAE_RAE_BENCH_1.0.text1K<n<10K1 likes290 downloads2y agoHugging Face06raei /Nieto-2022-ThinkingOutLoudOpenAccessEEGBasedBCIDatasetInnerSpeech Thinking out loud: an open-access EEG-based BCI dataset for inner speech recognition This is an unofficial mirror of OpenNeuro dataset ds003626, version 2.1.2. It is not affiliated with or endorsed by the dataset authors, their institutions, or OpenNeuro. Source and documentation Original dataset: OpenNeuro ds003626 v2.1.2 Article: Nieto et al., Scientific Data (2022) Original dataset documentation: README Official analysis code: N-Nieto/Inner_Speech_Dataset The… See the full description on the dataset page: https://huggingface.co/datasets/raei/Nieto-2022-ThinkingOutLoudOpenAccessEEGBasedBCIDatasetInnerSpeech.textn<1K0 likes254 downloads23d agoHugging Face07raeidsaqur /nifty-rl The News-Informed Financial Trend Yield (NIFTY) Dataset. The News-Informed Financial Trend Yield (NIFTY) Dataset. Details of the dataset, including data procurement and filtering can be found in the paper here: https://arxiv.org/abs/2405.09747. 📋 Table of Contents 🧩 NIFTY Dataset 📋 Table of Contents 📖 Usage Downloading the dataset Dataset structure Large Language Models ✍️ Contributing 📝 Citing 🙏 Acknowledgements 📖 Usage Downloading and using… See the full description on the dataset page: https://huggingface.co/datasets/raeidsaqur/nifty-rl.textmultiple-choice1K<n<10K1 likes132 downloads2y agoHugging Face08raeidsaqur /Hansard Pedagogical Machine Translation (Dialect) dataset: the filtered Canadian Hansard Dataset. The Canadian [Hansard](https://www.ourcommons.ca/documentviewer/en/35-2/house/hansard-index) is an archive of parliamentary sessions in the two official languages in Canada - English and Franch. 📋 Table of Contents 🧩 Hansard Dataset 📋 Table of Contents 📖 Usage Downloading the dataset Dataset structure Loading the dataset Loading the dataset The three partitions… See the full description on the dataset page: https://huggingface.co/datasets/raeidsaqur/Hansard.texttranslation100K<n<1M1 likes117 downloads3y agoHugging Face09HAERAE-HUB /HAE-RAE-COT-1.5M Dataset Card for "HAE-RAE-COT-1.5M" HAE-RAE-COT-1.5M is a dataset encompassing 1,586,688 samples of questions paired with CoT (Chain of Thought) rationales. The majority of this dataset is a translation of samples from the CoT-Collection, with a portion of samples derived from Korean datasets through the utilization of the gpt-3.5-turbo API. The translation of the CoT-Collection was carried out using the NLLB 600M model. To the best of our knowledge, HAE-RAE-COT-1.5M represents the… See the full description on the dataset page: https://huggingface.co/datasets/HAERAE-HUB/HAE-RAE-COT-1.5M.text1M<n<10M6 likes73 downloads3y agoHugging Face10RAED1ax /ChatGPT-Jailbreak-Prompts Dataset Card for Dataset Name Name ChatGPT Jailbreak Prompts Dataset Summary ChatGPT Jailbreak Prompts is a complete collection of jailbreak related prompts for ChatGPT. This dataset is intended to provide a valuable resource for understanding and generating text in the context of jailbreaking in ChatGPT. Languages [English] tabularquestion-answeringn<1K1 likes62 downloads8mo agoHugging Face11HAERAE-HUB /HAE_RAE_BENCH_2.0HAE_RAE_BENCH 2.0 is a miny implementation of Big-Bench consisted of 5 tasks: date_understanding, context_definition_alignment, proverb_unscrambling, 2_digit_multiply, and 3_digit_subtract. Paper Coming Soon (probably). text1K<n<10K4 likes37 downloads2y agoHugging Face12Satori-reasoning /Satori_RL_data_with_RAEtext100K<n<1M0 likes31 downloads1y agoHugging Face13raefdd /articles_2024-10-19tabularn<1K0 likes25 downloads2y agoHugging Face14raei /Brunner-2008-BCICompetition2008GrazDataSetA BCI Competition IV Dataset 2a — Graz data set A This is an unofficial mirror of the dataset described by C. Brunner, R. Leeb, G. R. Müller-Putz, A. Schlögl, and G. Pfurtscheller (2008) at Graz University of Technology. Originally released as BCI Competition IV Dataset 2a (BCICIV2A), it is distributed by BNCI Horizon 2020 as 001-2014. This mirror is not affiliated with or endorsed by the dataset contributors, their institutions, BNCI Horizon 2020, or the competition organizers.… See the full description on the dataset page: https://huggingface.co/datasets/raei/Brunner-2008-BCICompetition2008GrazDataSetA.documentn<1K0 likes19 downloads1d agoHugging Face15raefdd /articles_2024-10-21tabularn<1K0 likes14 downloads2y agoHugging Face16raefdd /article_inferences_2024-10-19textn<1K0 likes13 downloads2y agoHugging Face17NewEden-Forge /Rae-Taylor-Lora-Dataimagen<1K0 likes13 downloads2y agoHugging Face18raefdd /article_inferences_2024-10-20textn<1K0 likes12 downloads2y agoHugging Face19raefdd /articles_2024-10-22tabularn<1K0 likes8 downloads2y agoHugging Face20raefdd /articles_2024-10-23tabularn<1K0 likes8 downloads2y agoHugging Face21raefdd /article_inferences_2024-10-17textn<1K0 likes7 downloads2y agoHugging Face22raefdd /article_inferences_2024-10-24textn<1K0 likes7 downloads2y agoHugging Face23raekang /custom_llama_kor Dataset Card for "custom_llama_kor" More Information needed textn<1K0 likes6 downloads3y agoHugging Face24raefdd /article_inferences_2024-10-18textn<1K0 likes6 downloads2y agoHugging Face25raefdd /articles_2024-10-24tabularn<1K0 likes6 downloads2y agoHugging Face26raeemenon /en_hi_translationtext1K<n<10K0 likes6 downloads2y agoHugging Face27raefdd /articles_inferredtextn<1K0 likes5 downloads2y agoHugging Face28raefdd /articles_2024-10-16tabularn<1K0 likes5 downloads2y agoHugging Face29raefdd /article_inferences_2024-10-16textn<1K0 likes5 downloads2y agoHugging Face30raefdd /articles_2024-10-17tabularn<1K0 likes5 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.