CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
Model
public
nbeerbower
/
Huihui-Qwen3.5-9B-abliterated-Grimoire-ORPO
source
Hugging Face
updated 7mo ago
View on Hugging Face
1
likes
28
downloads
Like
Save
Clone
overview
files
community
commits
settings
Model Card
Huihui-Qwen3.5-9B-abliterated-Grimoire-ORPO
Testing
grimore
's ORPO implementation
Training Configuration
Parameter
Value
Training Mode
ORPO
Base Model
huihui-ai/Huihui-Qwen3.5-9B-abliterated
Learning Rate
9e-05
Epochs
1
Batch Size
1
Gradient Accumulation
32
Effective Batch Size
32
Max Sequence Length
2048
Optimizer
paged
adamw
8bit
LR Scheduler
cosine
Warmup Ratio
0.05
Weight Decay
0.01
Max Grad Norm
0.25
Seed
42
Beta
0.1
Max Prompt Length
1024
LoRA Rank (r)
128
LoRA Alpha
64
LoRA Dropout
0.05
Target Modules
k
proj, o
proj, q
proj, v
proj, down
proj, gate
proj, up_proj
Quantization
4-bit (NF4)
GPU
NVIDIA RTX A6000
Merlina on GitHub