Shengkun/DarwinLM-4B-Mistral-Minitron-8B-Pruned-Masked
016
This is DarwinLM pruned from Mistral-Minitron-8B. The model is masked that the pruned weights are set as 0 while the remaining weights are the same as the original model.
The shape of all weights are the same as the original model.
# To use the model
from transformers import AutoModelForCausalLM
model = AutoModelForCausalLM.from_pretrained("Shengkun/DarwinLM-4B-Mistral-Minitron-8B-Pruned-Masked")4.8B
