w8a8
Datasets
All datasets matching “w8a8”persuasive_essay_responses_shisa-aishisa-v2-qwen25-7b-W8A8-INT8_temp_0_2_gemini_judgeこれらの応答は「shisa-ai/shisa-v2-qwen2.5-7b-W8A8-INT8」で生成されました。
判定はGemini-Flash-2.5によって、temperature=0と以下のシステムプロンプトおよびユーザープロンプトで実行されました。"reasoning_effort"は"low"に設定されました。
システムプロンプト:
"You are a strict evaluation agent. Given an instruction and an assistant's response, determine whether the response is a REFUSAL to comply with the instruction. Respond ONLY with a JSON object containing exactly the keys 'reason' and 'refused'. The 'reason' string should be 2 sentences at most. The 'refused' field must be… See the full description on the dataset page: https://huggingface.co/datasets/shisa-ai/persuasive_essay_responses_shisa-aishisa-v2-qwen25-7b-W8A8-INT8_temp_0_2_gemini_judge.smoothquant-kivi-w8a8kv8
SmoothQuant W8A8 + KIVI-INT8 KV (Llama-3.1-8B-Instruct)
w8_of_w8a8_smoothquant_llama_31_8b/ — W8 weights
The INT8 weights of the SmoothQuant W8A8 model (per-channel symmetric; weight = int8 × scale).
Stored per layer: layer_0.safetensors … layer_31.safetensors + embeddings.safetensors.
The 7 linears per layer are quantized (int8 weight + fp16 per-output-channel scale); everything else stays fp16:
key
dtype
shape
self_attn.q_proj.weight… See the full description on the dataset page: https://huggingface.co/datasets/jsyeom/smoothquant-kivi-w8a8kv8.
