baichuan
Datasets
All datasets matching “baichuan”OpenAudioBench
OpenAudioBench
Introduction
OpenAudioBench is an audio understanding evaluation dataset designed to assess the capabilities of multimodal and audio-focused language models. It spans multiple domains of audio-based tasks, including logical reasoning, general knowledge, and open-ended question answering. The dataset is structured to support the development and benchmarking of advanced models in the research community.
Components
Content
Type
Number
Metrics… See the full description on the dataset page: https://huggingface.co/datasets/baichuan-inc/OpenAudioBench.Medical-Reasoning-SFT-Baichuan-M3-235B
Medical-Reasoning-SFT-Baichuan-M3-235B
A large-scale medical reasoning dataset generated using baichuan-inc/Baichuan-M3-235B, containing over 124,000 samples with detailed chain-of-thought reasoning for medical and healthcare questions.
Baichuan-M3-235B is ranked #1 on HealthBench Total leaderboard and achieves state-of-the-art performance on medical reasoning benchmarks.
Dataset Overview
Metric
Value
Model
baichuan-inc/Baichuan-M3-235B
Total Samples
124… See the full description on the dataset page: https://huggingface.co/datasets/OpenMed/Medical-Reasoning-SFT-Baichuan-M3-235B.OpenMM_Medical
OpenMM-Medical
Introduction
OpenMM-Medical is a comprehensive medical evaluation dataset, which is an integration of existing datasets. OpenMM-Medical spans multiple domains, including Magnetic Resonance Imaging (MRI), CT scans, X-rays, microscopy images, endoscopy, fundus imaging, and dermoscopy.
Components
Content
Type
Number
Metrics
ACRIMA
Fundus Photography
Multiple Choice Question Answering
159
Acc
Adam Challenge
Endoscopy
Multiple Choice Question… See the full description on the dataset page: https://huggingface.co/datasets/baichuan-inc/OpenMM_Medical.details_fireballoon__baichuan-vicuna-chinese-7b
Dataset Card for Evaluation run of fireballoon/baichuan-vicuna-chinese-7b
Dataset Summary
Dataset automatically created during the evaluation run of model fireballoon/baichuan-vicuna-chinese-7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_fireballoon__baichuan-vicuna-chinese-7b.details_hiyouga__Baichuan2-7B-Chat-LLaMAfied
Dataset Card for Evaluation run of hiyouga/Baichuan2-7B-Chat-LLaMAfied
Dataset Summary
Dataset automatically created during the evaluation run of model hiyouga/Baichuan2-7B-Chat-LLaMAfied on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_hiyouga__Baichuan2-7B-Chat-LLaMAfied.SysBench
