CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
Model
public
nbeerbower
/
Huihui-Qwen3.5-9B-abliterated-Grimoire-SFT
source
Hugging Face
updated 6mo ago
View on Hugging Face
1
likes
25
downloads
Like
Save
Clone
overview
files
community
commits
settings
Model Card
Huihui-Qwen3.5-9B-abliterated-Grimoire-SFT
Testing
grimore
's SFT implementation
Training Configuration
Parameter
Value
Training Mode
SFT
Base Model
huihui-ai/Huihui-Qwen3.5-9B-abliterated
Learning Rate
9e-05
Epochs
1
Batch Size
2
Gradient Accumulation
16
Effective Batch Size
32
Max Sequence Length
2048
Optimizer
paged
adamw
8bit
LR Scheduler
cosine
Warmup Ratio
0.05
Weight Decay
0.01
Max Grad Norm
0.25
Seed
42
LoRA Rank (r)
128
LoRA Alpha
64
LoRA Dropout
0.05
Target Modules
up
proj, down
proj, gate
proj, k
proj, q
proj, v
proj, o_proj
Quantization
4-bit (NF4)
GPU
NVIDIA RTX A6000
Merlina on GitHub