zlaabsi/Qwen3.6-27B-OTQ-GGUF-benchmarks
Qwen3.6-27B OTQ GGUF Benchmark Reproducibility This dataset contains the compact paired benchmark evidence used by zlaabsi/Qwen3.6-27B-OTQ-GGUF. It is a reproducibility dataset, not a leaderboard dataset. The rows are small practical release signals run on pinned task IDs with prompt format qwen3-no-think, deterministic decoding and local scoring rules. Contents Path Meaning data/paired_samples.jsonl Flattened 232-row paired sample table with prompts… See the full description on the dataset page: https://huggingface.co/datasets/zlaabsi/Qwen3.6-27B-OTQ-GGUF-benchmarks.
Qwen3.6-27B OTQ GGUF Benchmark Reproducibility
This dataset contains the compact paired benchmark evidence used by `zlaabsi/Qwen3.6-27B-OTQ-GGUF`.
It is a reproducibility dataset, not a leaderboard dataset. The rows are small practical release signals run on pinned task IDs with prompt format qwen3-no-think, deterministic decoding and local scoring rules.
Contents
Scope
- Base model:
Qwen/Qwen3.6-27B - GGUF model repo:
zlaabsi/Qwen3.6-27B-OTQ-GGUF - BF16 runtime: Hugging Face Jobs H200 with Transformers
- GGUF runtime: stock
llama.cpp/llama-server, Metal, FlashAttention - Prompt format:
qwen3-no-think - Sample count: 232
These files include prompts derived from upstream benchmark datasets. Respect the upstream dataset licenses and terms for any reuse beyond reproducibility of this release evaluation.
