CoolFace
20 results

baichuan

baichuan-inc /OpenAudioBench OpenAudioBench Introduction OpenAudioBench is an audio understanding evaluation dataset designed to assess the capabilities of multimodal and audio-focused language models. It spans multiple domains of audio-based tasks, including logical reasoning, general knowledge, and open-ended question answering. The dataset is structured to support the development and benchmarking of advanced models in the research community. Components Content Type Number Metrics… See the full description on the dataset page: https://huggingface.co/datasets/baichuan-inc/OpenAudioBench.audio1K<n<10K8 likes1.1k downloads2y agoHugging FaceOpenMed /Medical-Reasoning-SFT-Baichuan-M3-235B Medical-Reasoning-SFT-Baichuan-M3-235B A large-scale medical reasoning dataset generated using baichuan-inc/Baichuan-M3-235B, containing over 124,000 samples with detailed chain-of-thought reasoning for medical and healthcare questions. Baichuan-M3-235B is ranked #1 on HealthBench Total leaderboard and achieves state-of-the-art performance on medical reasoning benchmarks. Dataset Overview Metric Value Model baichuan-inc/Baichuan-M3-235B Total Samples 124… See the full description on the dataset page: https://huggingface.co/datasets/OpenMed/Medical-Reasoning-SFT-Baichuan-M3-235B.texttext-generation100K<n<1M7 likes173 downloads8mo agoHugging Facebaichuan-inc /OpenMM_Medical OpenMM-Medical Introduction OpenMM-Medical is a comprehensive medical evaluation dataset, which is an integration of existing datasets. OpenMM-Medical spans multiple domains, including Magnetic Resonance Imaging (MRI), CT scans, X-rays, microscopy images, endoscopy, fundus imaging, and dermoscopy. Components Content Type Number Metrics ACRIMA Fundus Photography Multiple Choice Question Answering 159 Acc Adam Challenge Endoscopy Multiple Choice Question… See the full description on the dataset page: https://huggingface.co/datasets/baichuan-inc/OpenMM_Medical.image5 likes168 downloads2y agoHugging Faceopen-llm-leaderboard-old /details_fireballoon__baichuan-vicuna-chinese-7b Dataset Card for Evaluation run of fireballoon/baichuan-vicuna-chinese-7b Dataset Summary Dataset automatically created during the evaluation run of model fireballoon/baichuan-vicuna-chinese-7b on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_fireballoon__baichuan-vicuna-chinese-7b.0 likes167 downloads3y agoHugging Faceopen-llm-leaderboard-old /details_hiyouga__Baichuan2-7B-Chat-LLaMAfied Dataset Card for Evaluation run of hiyouga/Baichuan2-7B-Chat-LLaMAfied Dataset Summary Dataset automatically created during the evaluation run of model hiyouga/Baichuan2-7B-Chat-LLaMAfied on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_hiyouga__Baichuan2-7B-Chat-LLaMAfied.0 likes62 downloads3y agoHugging FacePKU-Baichuan-MLSystemLab /SysBench2 likes45 downloads2y agoHugging Face