CoolFace
Modelpublic

agentlans/Llama3.1-SuperDeepFuse

sourceHugging Facellama3.1updated 2y agoView on Hugging Face
1likes33downloads
Model Card

Llama3.1-SuperDeepFuse

An 8B parameter language model that merges three high-performance distilled models to boost reasoning, instruction-following, and performance in mathematics and coding.

Model Highlights

Key Capabilities

  • —Enhanced multi-task reasoning
  • —Improved mathematical and coding performance
  • —Multilingual support

Performance Notes

  • —Maintains Llama 3.1 safety standards
  • —Suitable for consumer GPU deployment
  • —Balanced performance across diverse tasks

Considerations

  • —Still being benchmarked
  • —Capabilities limited compared to larger model variants
  • —Can give misleading output like all other language models
  • —Outputs should be independently verified

Licensing

Follows standard Llama 3.1 usage terms.

Open LLM Leaderboard Evaluation Results

Detailed results can be found here! Summarized results can be found here!

MetricValue (%)
Average27.30
IFEval (0-Shot)77.62
BBH (3-Shot)29.22
MATH Lvl 5 (4-Shot)17.75
GPQA (0-shot)3.24
MuSR (0-shot)5.13
MMLU-PRO (5-shot)30.83