CoolFace
20 results

tcm

JX-Lab /TCM-MKG TCM-MKG Reorganized Graph Export 本仓库是对原始 TCM-MKG V1.0 的格式整理与图结构导出,不是原始数据集的官方发布,也不声称拥有底层知识内容。 Dataset Description This repository contains a Hugging Face-compatible, reorganized graph export derived from the publicly available Traditional Chinese Medicine Multidimensional Knowledge Graph (TCM-MKG V1.0). JX-Lab performed format conversion and schema normalization to make the resource easier to load for knowledge graph and machine learning workflows. The repository… See the full description on the dataset page: https://huggingface.co/datasets/JX-Lab/TCM-MKG.textother100K<n<1M0 likes1.6k downloads17d agoHugging FaceCarsonnnNN /TCM-Pretrain-Data-ShizhenGPT 📚 Introduction This dataset is the pre-training dataset for ShizhenGPT, a multimodal LLM for Traditional Chinese Medicine (TCM). We open-source the largest existing TCM corpus dataset (over 5B tokens) from TCM-related websites and books. Additionally, we also open-source the largest scale TCM image-text pretraining dataset. For details, see our paper and GitHub repository. 📊 Dataset Overview The open-sourced pre-training dataset consists of five parts:… See the full description on the dataset page: https://huggingface.co/datasets/CarsonnnNN/TCM-Pretrain-Data-ShizhenGPT.texttext-generation1M<n<10M1 likes1.3k downloads7mo agoHugging Facealirezafalah /ELTE-TCM-46k CNC Tool Condition Image Dataset (ELTE-TCM-46k) This dataset contains approximately 46,000 high-resolution TIFF images of 116 unique CNC cutting tools (including drills, end mills, and chamfer tools) under various conditions. The images were captured to support research in automated, direct Tool Condition Monitoring (TCM). The primary purpose of this dataset is to facilitate the novel methodology presented in Falah, Andó, & Szekeres (2025), which transforms a sequence of 2D… See the full description on the dataset page: https://huggingface.co/datasets/alirezafalah/ELTE-TCM-46k.imageimage-classificationn<1K1 likes934 downloads4d agoHugging FaceFreedomIntelligence /TCM-Pretrain-Data-ShizhenGPT 📚 Introduction This dataset is the pre-training dataset for ShizhenGPT, a multimodal LLM for Traditional Chinese Medicine (TCM). We open-source the largest existing TCM corpus dataset (over 5B tokens) from TCM-related websites and books. Additionally, we also open-source the largest scale TCM image-text pretraining dataset. For details, see our paper and GitHub repository. 📊 Dataset Overview The open-sourced pre-training dataset consists of five parts:… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/TCM-Pretrain-Data-ShizhenGPT.texttext-generation1M<n<10M10 likes672 downloads1y agoHugging FaceZJUFanLab /TCMChat-dataset-600k中文 | English TCMChat: A Generative Large Language Model for Traditional Chinese Medicine News [2024-11-1] We have fully open-sourced the model weights and training dataset on Huggingface.[2024-5-17] Open source model weight on HuggingFace. Application Install git clone https://github.com/ZJUFanLab/TCMChat cd TCMChat Create a conda environment conda create -n baichuan2 python=3.10 -y First install the dependency package.… See the full description on the dataset page: https://huggingface.co/datasets/ZJUFanLab/TCMChat-dataset-600k.15 likes658 downloads2y agoHugging Facetcm03 /SnapUGC_1video10K<n<100K0 likes480 downloads1y agoHugging Face