NexesMess/Llama_3.x_70b_Hexagon_Blue_V2
06
about
On testing for now. V1 seems to be smarter, more focused. The Nemotricks merge blurs things, it seems. I'll keep this V2 in the Mess.
merge
This is a merge of pre-trained language models created using mergekit.
Merge Details
Merge Method
This model was merged using the Model Stock merge method using huihui-ai/Llama-3.3-70B-Instruct-abliterated as a base.
Models Merged
The following models were included in the merge:
- Nexesenex/Llama_3.3_70b_DarkHorse
- huihui-ai/Llama-3.1-Tulu-3-70B-abliterated
- TheDrummer/Fallen-Llama-3.3-R1-70B-v1
- hitachi-nlp/Llama-3.1-70B-FLDx2
- Nexesenex/Llama_3.1_70b_Nemotricks_v1.0
Configuration
The following YAML configuration was used to produce this model:
merge_method: model_stock
models:
- model: TheDrummer/Fallen-Llama-3.3-R1-70B-v1
parameters:
weight: 1.0
- model: Nexesenex/Llama_3.3_70b_DarkHorse
parameters:
weight: 1.0
- model: Nexesenex/Llama_3.1_70b_Nemotricks_v1.0
parameters:
weight: 1.0
- model: huihui-ai/Llama-3.1-Tulu-3-70B-abliterated
parameters:
weight: 1.0
- model: hitachi-nlp/Llama-3.1-70B-FLDx2
parameters:
weight: 1.0
base_model: huihui-ai/Llama-3.3-70B-Instruct-abliterated
dtype: bfloat16
out_dtype: bfloat16
parameters:
int8_mask: true
normalize: true
rescale: false
filter_wise: false
smooth: false
allow_negative_weights: false
chat_template: auto
tokenizer:
source: union