CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01dharmajitbaro /ham10kimage1K<n<10K1 likes555 downloads2mo agoHugging Face02Dharma-AI /DharmaOCR-Benchmark DharmaOCR-Benchmark Overview DharmaOCR-Benchmark is a 496-instance evaluation suite for OCR models focused on Brazilian Portuguese documents. It covers printed text, handwritten text, and legal/administrative documents — domains underrepresented in existing benchmarks like OCRBench and olmOCR-Bench. This benchmark evaluates not only transcription quality, but also text degeneration rate and unit inference cost as first-class metrics. Released alongside the… See the full description on the dataset page: https://huggingface.co/datasets/Dharma-AI/DharmaOCR-Benchmark.imageimage-text-to-textn<1K6 likes456 downloads2mo agoHugging Face03garydean /defining-dharma in search of dharma — research corpus Dharma here means how human societies build, transmit and enforce ethics — not metaphysics. A traceably-sourced research corpus: every claim cited, drawn from anthropology, evolutionary biology, history and comparative ethics. Written by Gary Dean (Biksu Okusi). The corpus has three configs serving different purposes: notes (52 records, 1562 cited sources) — Stage-1 research notes. Each note answers one registry question and carries… See the full description on the dataset page: https://huggingface.co/datasets/garydean/defining-dharma.texttext-generationn<1K1 likes169 downloads12m agoHugging Face04pharaouk /dharma-1 "Dharma-1" A new carefully curated benchmark set, designed for a new era where the true end user uses LLM's for zero-shot and one-shot tasks, for a vast majority of the time. Stop training your models on mindless targets (eval_loss, train_loss), start training your LLM on lightweight Dharma as an eval target. A mix of all the top benchmarks. Formed to have an equal distribution of some of the most trusted benchmarks used by those developing SOTA LLMs, comprised of only 3,000… See the full description on the dataset page: https://huggingface.co/datasets/pharaouk/dharma-1.text1K<n<10K41 likes99 downloads3y agoHugging Face05pharaouk /dharma-2 "dharma_g1i5 Dataset" A dharma evaluation dataset with the following configuration: ||| Subject: MMLU, Size: 38 ||| ||| Subject: ARC-Challenge, Size: 38 ||| ||| Subject: ARC-Easy, Size: 38 ||| ||| Subject: BoolQ, Size: 35 ||| ||| Subject: winogrande, Size: 38 ||| ||| Subject: openbookqa, Size: 38 ||| ||| Subject: truthful_qa, Size: 38 ||| ||| Subject: agieval, Size: 37 ||| Made with https://github.com/pharaouk/dharma 🚀 textn<1K0 likes42 downloads2y agoHugging Face06joyboseroy /bengal-dharma-corpus Bengal Dharma Corpus Evolution of Bengali Devotional Language: a multi-tradition corpus spanning Old Bengali, Sanskrit, and modern Bengali across Buddhist, Shakta, and Vaishnava traditions, 8th to 19th century. Assembled and curated by Joy Bose (joyboseroy), June 2026. Code and analysis: https://github.com/joyboseroy/bengal-dharma-corpus Related dataset: joyboseroy/darshana-graph (arXiv:2606.18222) What this corpus is This is a curated collection of 75 texts from… See the full description on the dataset page: https://huggingface.co/datasets/joyboseroy/bengal-dharma-corpus.tabularn<1K0 likes37 downloads3mo agoHugging Face07munmuni /DharmaBench_v1 DharmaBench v1 DharmaBench v1 is a Bengali spiritual Question-Answering dataset. It contains teachings and quotes from Sri Ramakrishna, Sri Sarada Devi, Swami Vivekananda, and other holy personalities. Dataset Description This dataset is created to build AI models that can answer people's life problems with spiritual wisdom from Bengali sources. Dataset Structure The dataset has 50 examples with the following columns: id: Unique ID for each… See the full description on the dataset page: https://huggingface.co/datasets/munmuni/DharmaBench_v1.textquestion-answeringn<1K0 likes36 downloads2mo agoHugging Face08dharmam-stjude /CT-RATE-Dataset-cleanedtabular10K<n<100K0 likes31 downloads10mo agoHugging Face09thewulf7 /dharma-catalogtextn<1K0 likes30 downloads16d agoHugging Face10manishiitg /pharaouk_dharma-1-hitext1K<n<10K0 likes23 downloads3y agoHugging Face11pharaouk /dharma_test3 "dharma_test2 Dataset" A dharma evaluation dataset with the following configuration: ||| Subject: MMLU, Size: 12 ||| ||| Subject: ARC-Challenge, Size: 12 ||| ||| Subject: ARC-Easy, Size: 12 ||| ||| Subject: BoolQ, Size: 12 ||| ||| Subject: winogrande, Size: 12 ||| ||| Subject: openbookqa, Size: 12 ||| ||| Subject: truthful_qa, Size: 12 ||| ||| Subject: agieval, Size: 12 ||| Made with https://github.com/pharaouk/dharma 🚀 0 likes22 downloads3y agoHugging Face12pharaouk /dharma-test "dharma-test Dataset" A dharma evaluation dataset with the following configuration: Subject: MMLU, Size: 12 Subject: ARC-Challenge, Size: 12 Subject: ARC-Easy, Size: 12 Subject: BoolQ, Size: 12 Subject: winogrande, Size: 12 Subject: openbookqa, Size: 12 Subject: truthful_qa, Size: 12 Subject: agieval, Size: 12 Made with https://github.com/pharaouk/dharma 🚀 0 likes19 downloads3y agoHugging Face13pharaouk /dharma_test2 "dharma_test2 Dataset" A dharma evaluation dataset with the following configuration: ||| Subject: MMLU, Size: 12 ||| ||| Subject: ARC-Challenge, Size: 12 ||| ||| Subject: ARC-Easy, Size: 12 ||| ||| Subject: BoolQ, Size: 12 ||| ||| Subject: winogrande, Size: 12 ||| ||| Subject: openbookqa, Size: 12 ||| ||| Subject: truthful_qa, Size: 12 ||| ||| Subject: agieval, Size: 12 ||| Made with https://github.com/pharaouk/dharma 🚀 0 likes19 downloads3y agoHugging Face14pharaouk /dharma-test2 "dharma-test2 Dataset" A dharma evaluation dataset with the following configuration: ||| Subject: MMLU, Size: 12 ||| ||| Subject: ARC-Challenge, Size: 12 ||| ||| Subject: ARC-Easy, Size: 12 ||| ||| Subject: BoolQ, Size: 12 ||| ||| Subject: winogrande, Size: 12 ||| ||| Subject: openbookqa, Size: 12 ||| ||| Subject: truthful_qa, Size: 12 ||| ||| Subject: agieval, Size: 12 ||| Made with https://github.com/pharaouk/dharma 🚀 0 likes18 downloads3y agoHugging Face15UmerHA /dharma2Copy of https://huggingface.co/datasets/pharaouk/dharma-2. All credit to pharaouk! I'll use this fixed copy until https://huggingface.co/datasets/pharaouk/dharma-2/discussions/1 is solves. textn<1K0 likes15 downloads2y agoHugging Face16Adun /dharma-thai-0012 likes11 downloads3y agoHugging Face17kala185 /Hindu_Dharma0 likes10 downloads2y agoHugging Face18JuniorThap /Dharma-Kid-Chattextn<1K0 likes9 downloads1y agoHugging Face19madhavkotecha /DharmaQAtext1K<n<10K1 likes8 downloads1y agoHugging Face20Ishant57 /dharmabot-indextext10K<n<100K0 likes6 downloads3mo agoHugging Face21Vincero /Dataset_dharmadatadocumentn<1K0 likes5 downloads1y agoHugging Face22DharmaBytes /functiongemma-netboxtextn<1K0 likes4 downloads8mo agoHugging Face23somnan /dharma-thai-0010 likes3 downloads3y agoHugging Face24Adun /dharma-thai-002textn<1K0 likes3 downloads3y agoHugging Face25Alignment-Lab-AI /Dharma-2textn<1K0 likes3 downloads2y agoHugging Face26dharmads /simdatatextn<1K0 likes3 downloads2y agoHugging Face27dharmatejadhulipudi /genai0 likes2 downloads2y agoHugging Face28sherab-choephel /paramarthaviniscaya-dharmacakra Paramārthaviniścaya Dharmacakra | དོན་དམ་རྣམ་ངེས་ཆོས་འཁོར། Purpose | དམིགས་ཡུལ། This dataset preserves and makes accessible the Third Turning of the Wheel of Dharma, specifically the Zhentong (Gzhan stong) view of Tibetan Buddhism. It contains source texts, commentaries, and explanations on Buddha-Nature, Tathāgatagarbha, and the luminous, empty nature of mind as taught in the Jonang, Kagyu, Nyingma and yogacharamadhyamaka school traditions.… See the full description on the dataset page: https://huggingface.co/datasets/sherab-choephel/paramarthaviniscaya-dharmacakra.0 likes2 downloads5mo agoHugging Face29fransekman /dharma0 likes1 downloads1y agoHugging Face30Dharmalakh /20 likes1 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.