CoolFace
Modelpublic

angeliko/Qwen-2.5-7b-S1k-bnb-4bit

sourceHugging Faceupdated 5mo agoView on Hugging Face
0likes5downloads
Model Card

bunnycore/Qwen-2.5-7b-S1k (Quantized)

Description

This model is a quantized version of the original model `bunnycore/Qwen-2.5-7b-S1k`.

It's quantized using the BitsAndBytes library to 4-bit using the bnb-my-repo space.

Quantization Details

  • —Quantization Type: int4
  • —bnb_4bit_quant_type: nf4
  • —bnb_4bit_use_double_quant: True
  • —bnb_4bit_compute_dtype: bfloat16
  • —bnb_4bit_quant_storage: uint8

📄 Original Model Information

System Prompt

Think about the reasoning process in the mind first, then provide the answer. The reasoning process should detailed and should be wrapped within <think> </think> tags, then provide the answer after that, i.e., <think> reasoning process here </think> answer here.

Configuration

The following YAML configuration was used to produce this model:

yaml

base_model: bunnycore/Qwen-2.5-7B-Deep-Stock-v4+bunnycore/Qwen-2.5-7b-s1k-lora_model
dtype: bfloat16
merge_method: passthrough
models:
  - model: bunnycore/Qwen-2.5-7B-Deep-Stock-v4+bunnycore/Qwen-2.5-7b-s1k-lora_model
tokenizer_source: bunnycore/Qwen-2.5-7B-Deep-Stock-v4

Open LLM Leaderboard Evaluation Results

Detailed results can be found here

MetricValue
Avg.34.59
IFEval (0-Shot)71.62
BBH (3-Shot)36.69
MATH Lvl 5 (4-Shot)47.81
GPQA (0-shot)4.59
MuSR (0-shot)9.26
MMLU-PRO (5-shot)37.58