Green-Sky/nakodanei-Blue-Orchid-2x7b-GGUF-iMatrix
8120
llama.cpp conversion of https://huggingface.co/nakodanei/Blue-Orchid-2x7b/
except for f16 and q8_0, every quant is using the merge.imatrix
merge.imatrix is a merge of kalomaze-group_10_merged.172chunks.imatrix and wiki.train.400chunks.imatrix, which took ~10min + ~20min to calulate on my machine.
full wiki.train would have taken 10h
for more info on imatrix handling see https://github.com/ggerganov/llama.cpp/pull/5302
ppl (512 wiki.test, 300chunks)
Interesting observations
despite merge.imatrix being different from kalomaze-group_10_merged.172chunks.imatrix, they produce the exact same quantized iq3_xxs model file. (same hash, checked multiple times)
q5km has a lower perplexity with the imatrix. but that probably is caused by kalomaze-group10merged diverging enough from wiki.
