datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
lm-eval-results-yunconglong-Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B-private
Dataset Card for Evaluation run of yunconglong/Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B
Dataset automatically created during the evaluation run of model yunconglong/Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yunconglong-Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B-private.lm-eval-results-TomGrc-FusionNet_7Bx2_MoE_14B-private
Dataset Card for Evaluation run of TomGrc/FusionNet_7Bx2_MoE_14B
Dataset automatically created during the evaluation run of model TomGrc/FusionNet_7Bx2_MoE_14B
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-TomGrc-FusionNet_7Bx2_MoE_14B-private.lm-eval-results-TomGrc-FusionNet_7Bx2_MoE_v0.1-private
Dataset Card for Evaluation run of TomGrc/FusionNet_7Bx2_MoE_v0.1
Dataset automatically created during the evaluation run of model TomGrc/FusionNet_7Bx2_MoE_v0.1
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-TomGrc-FusionNet_7Bx2_MoE_v0.1-private.lm-eval-results-dddsaty-FusionNet_7Bx2_MoE_Ko_DPO_Adapter_Attach-private
Dataset Card for Evaluation run of dddsaty/FusionNet_7Bx2_MoE_Ko_DPO_Adapter_Attach
Dataset automatically created during the evaluation run of model dddsaty/FusionNet_7Bx2_MoE_Ko_DPO_Adapter_Attach
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-dddsaty-FusionNet_7Bx2_MoE_Ko_DPO_Adapter_Attach-private.sample-fusion-intelligence-traces
Sample Fusion Intelligence Traces
Structured AI reasoning traces from dFusion's Fusion Intelligence system. Each record captures a complete agentic workflow: a real user query on a domain-specific topic, the full message chain including system prompts, tool calls, search results, intermediate reasoning steps, and a final synthesized answer — along with human feedback.
These are not synthetic benchmarks. They are traces from real queries submitted by real users on live financial… See the full description on the dataset page: https://huggingface.co/datasets/dFusionAILabs/sample-fusion-intelligence-traces.douvras-environmental-sensor-fusion
Douvras Environmental Sensor Fusion v0.1
Synthetic episodes combining air quality, noise, traffic, light and temperature.
Labels include normal operation, single-sensor spikes, multi-sensor anomaly and
missing sensor. It contains 72 records (48/12/12) across 12 episodes, split by
episode. No real sensor or location data is included.
This is a fusion protocol, not an operational alarm system. Human review is
required before any intervention.
sthenno__tempesthenno-fusion-0309-details
Dataset Card for Evaluation run of sthenno/tempesthenno-fusion-0309
Dataset automatically created during the evaluation run of model sthenno/tempesthenno-fusion-0309
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sthenno__tempesthenno-fusion-0309-details.bunnycore__Qwen2.5-7B-Instruct-Fusion-details
Dataset Card for Evaluation run of bunnycore/Qwen2.5-7B-Instruct-Fusion
Dataset automatically created during the evaluation run of model bunnycore/Qwen2.5-7B-Instruct-Fusion
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/bunnycore__Qwen2.5-7B-Instruct-Fusion-details.sometimesanotion__Lamarck-14B-v0.7-Fusion-details
Dataset Card for Evaluation run of sometimesanotion/Lamarck-14B-v0.7-Fusion
Dataset automatically created during the evaluation run of model sometimesanotion/Lamarck-14B-v0.7-Fusion
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/sometimesanotion__Lamarck-14B-v0.7-Fusion-details.FINGU-AI__Chocolatine-Fusion-14B-details
Dataset Card for Evaluation run of FINGU-AI/Chocolatine-Fusion-14B
Dataset automatically created during the evaluation run of model FINGU-AI/Chocolatine-Fusion-14B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/FINGU-AI__Chocolatine-Fusion-14B-details.qingy2024__Fusion-14B-Instruct-details
Dataset Card for Evaluation run of qingy2024/Fusion-14B-Instruct
Dataset automatically created during the evaluation run of model qingy2024/Fusion-14B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/qingy2024__Fusion-14B-Instruct-details.qingy2024__Fusion2-14B-Instruct-details
Dataset Card for Evaluation run of qingy2024/Fusion2-14B-Instruct
Dataset automatically created during the evaluation run of model qingy2024/Fusion2-14B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/qingy2024__Fusion2-14B-Instruct-details.jaspionjader__Kosmos-EVAA-Fusion-8B-details
Dataset Card for Evaluation run of jaspionjader/Kosmos-EVAA-Fusion-8B
Dataset automatically created during the evaluation run of model jaspionjader/Kosmos-EVAA-Fusion-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/jaspionjader__Kosmos-EVAA-Fusion-8B-details.huihui-ai__QwQ-32B-Coder-Fusion-8020-details
Dataset Card for Evaluation run of huihui-ai/QwQ-32B-Coder-Fusion-8020
Dataset automatically created during the evaluation run of model huihui-ai/QwQ-32B-Coder-Fusion-8020
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/huihui-ai__QwQ-32B-Coder-Fusion-8020-details.huihui-ai__QwQ-32B-Coder-Fusion-7030-details
Dataset Card for Evaluation run of huihui-ai/QwQ-32B-Coder-Fusion-7030
Dataset automatically created during the evaluation run of model huihui-ai/QwQ-32B-Coder-Fusion-7030
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/huihui-ai__QwQ-32B-Coder-Fusion-7030-details.qingy2024__Fusion4-14B-Instruct-details
Dataset Card for Evaluation run of qingy2024/Fusion4-14B-Instruct
Dataset automatically created during the evaluation run of model qingy2024/Fusion4-14B-Instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/qingy2024__Fusion4-14B-Instruct-details.OmnicromsBrain__NeuralStar_FusionWriter_4x7b-details
Dataset Card for Evaluation run of OmnicromsBrain/NeuralStar_FusionWriter_4x7b
Dataset automatically created during the evaluation run of model OmnicromsBrain/NeuralStar_FusionWriter_4x7b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/OmnicromsBrain__NeuralStar_FusionWriter_4x7b-details.localization-sensor-fusion-json
Localization Sensor Fusion Dataset
JSON dataset combining multiple sensor readings
for robot localization tasks.
