CoolFace
Modelpublic

LevArtesa/grpo-humanizer-de

sourceHugging Faceapache-2.0updated 5mo agoView on Hugging Face
0likes28downloads
Model Card

GRPO Humanizer DE

Fine-tuned with Group Relative Policy Optimization (GRPO) to rewrite AI-generated German academic text so that it passes GPTZero detection while preserving semantic content.

Training details

ParameterValue
Base modelQwen/Qwen3-8B
MethodGRPO (TRL) + LoRA
Learning rate5e-06
Batch size2
Gradient accumulation8
Max steps50
Precisionbf16

Intended use

Academic text humanisation for German-language content. The model is designed to be called via the HuggingFace Inference API from the GhostWriter application.

Licence

Apache-2.0