CoolFace
Modelpublic

Lordnyx/qwen3loop-0.6b-sft-deep-supervision-v1

sourceHugging Faceapache-2.0updated 2d agoView on Hugging Face
1likes3.1kdownloads
50 commits on main
ae27a4a2d ago

Update modelo_qwen3loop_sft_f16.gguf: modelo_qwen3loop_sft_f16.gguf (28 blocos nativos, FP16 - 1.20 GB)

Lordnyx
d1c72722d ago

Update unrolled_modelo_qwen3loop_sft_f16.gguf: unrolled_modelo_qwen3loop_sft_f16.gguf (56 camadas, FP16 - 2.08 GB)

Lordnyx
54a154f2d ago

Update modelo_qwen3loop_sft_q8_0.gguf: modelo_qwen3loop_sft_q8_0.gguf (28 blocos nativos, Q8_0 - 0.64 GB)

Lordnyx
dc660b12d ago

Update unrolled_modelo_qwen3loop_sft_q8_0.gguf: unrolled_modelo_qwen3loop_sft_q8_0.gguf (56 camadas, Q8_0 - 1.11 GB)

Lordnyx
b2953cd2d ago

Update model.safetensors: model.safetensors (Pesos finais do SFT - 1.14 GB)

Lordnyx
d07f20c2d ago

Update latent_halting_probe.pt: Sonda MLP Latente (latent_halting_probe.pt - 91.5% acc)

Lordnyx
9553fd32d ago

Update engine/qwen3loop_python/modeling_qwen3loop.py: modeling_qwen3loop.py com suporte corrigido a loop_hidden_states

Lordnyx
63eb8d52d ago

Update config.json: config.json atualizado

Lordnyx
c14f9742d ago

Update README.md: Model Card (README.md) atualizado com métricas SFT e benchmarks

Lordnyx
d9445883d ago

Fix tokenizer tokens and merges in modelo_qwen3loop_sft_q8_0.gguf

Lordnyx
f73894f3d ago

Fix tokenizer tokens and merges in unrolled_modelo_qwen3loop_sft_q8_0.gguf

Lordnyx
65b1b693d ago

Fix tokenizer tokens and merges in unrolled_modelo_qwen3loop_sft_f16.gguf

Lordnyx
e9107b23d ago

Update unrolled_modelo_qwen3loop_sft_q8_0.gguf (hotfix 0.3 unrolled 56 layers q8_0)

Lordnyx
a572e593d ago

Update unrolled_modelo_qwen3loop_sft_f16.gguf (hotfix 0.3 unrolled 56 layers qwen3)

Lordnyx
b8c2c323d ago

Update modelo_qwen3loop_sft_q8_0.gguf (hotfix 0.3)

Lordnyx
1054aa33d ago

Update modelo_qwen3loop_sft_f16.gguf (hotfix 0.3)

Lordnyx
18a3abd3d ago

engine: update README for triple-loop patch (qwen3loopdelta superseded)

Lordnyx
50cc8f33d ago

engine: add qwen35 triple-loop llama.cpp patch (32->64 shared-weight loop)

Lordnyx
0231e2d3d ago

engine: add qwen35 triple-loop llama.cpp patch (32->64 shared-weight loop)

Lordnyx
e63f6773d ago

engine: add qwen35 triple-loop llama.cpp patch (32->64 shared-weight loop)

Lordnyx
7eda6ce3d ago

engine: remove superseded qwen3loopdelta llama.cpp patch (see git history to recover)

Lordnyx
b1b700c8d ago

model: Hotfix 0.3 - Update standalone model.safetensors with DS-SimPO anti-loop weights

Lordnyx
f116c828d ago

docs: Update README.md with Hotfix 0.3 benchmarks and long-horizon stability analysis

Lordnyx
9cdfc7b8d ago

docs: Add Hotfix 0.3 technical release notes (DS-SimPO anti-loop alignment)

Lordnyx
e18bb921mo ago

Hotfix 0.2: Add Unrolled universal model documentation, VRAM trade-off warning and parameter guide

Lordnyx
55ffe7c1mo ago

Add unrolled_modelo_qwen3loop_sft_q8_0.gguf: 42-layer standard Qwen3 Q8_0 binary for Ollama / LM Studio (Zero patches required)

Lordnyx
462a95e1mo ago

Add unrolled_modelo_qwen3loop_sft_f16.gguf: 42-layer standard Qwen3 architecture (Zero C++ patches required)

Lordnyx
2981d241mo ago

Hotfix 0.2: English technical release notes & playground readiness guidelines

Lordnyx
34233e31mo ago

Hotfix 0.2: Update README with playground readiness, sampling parameters and layerwise benchmark

Lordnyx
4bd109f1mo ago

Hotfix 0.2: Technical release notes, layer probing metrics & playground readiness parameters

Lordnyx
6cee45f1mo ago

Hotfix 0.2: Update modelo_qwen3loop_sft_q8_0.gguf with Suffix-Calibrated Step 75 quantized binary

Lordnyx
a9d7e241mo ago

Hotfix 0.2: Update modelo_qwen3loop_sft_f16.gguf with Suffix-Calibrated Step 75 production binary

Lordnyx
64df0661mo ago

Hotfix 0.2: Update model.safetensors with Calibrated Suffix weights (Step 75)

Lordnyx
0e628601mo ago

Hotfix 0.2: Update generation_config.json with Calibrated Suffix weights (Step 75)

Lordnyx
fc98bc01mo ago

Hotfix 0.2: Update config.json with Calibrated Suffix weights (Step 75)

Lordnyx
0ddae8d1mo ago

Hotfix 0.1: Add release notes, layer metrics, and benchmark telemetry (hotfix0.1.txt)

Lordnyx
8a252811mo ago

Hotfix 0.1: Update modelo_qwen3loop_sft_q8_0.gguf with DARE-TIES production quantized binary

Lordnyx
2da6c201mo ago

Hotfix 0.1: Update modelo_qwen3loop_sft_f16.gguf with DARE-TIES best checkpoint

Lordnyx
52fbd951mo ago

Hotfix 0.1: Update tokenizer_config.json with DARE-TIES global weights

Lordnyx
15934dd1mo ago

Hotfix 0.1: Update model.safetensors with DARE-TIES global weights

Lordnyx
b8b4ed81mo ago

Hotfix 0.1: Update config.json with DARE-TIES global weights

Lordnyx
9ea3d4b1mo ago

Upload README.md with huggingface_hub

Lordnyx
ef745491mo ago

Publish qwen3loop custom engine (Python package + llama.cpp patch)

Lordnyx
88dddc31mo ago

Upload README.md with huggingface_hub

Lordnyx
4f132991mo ago

Upload README.md with huggingface_hub

Lordnyx
a35fa501mo ago

Update README with credits and dataset link to ianncity/GLM-5.2-Logic-Puzzles

Lordnyx
33d61871mo ago

Upload modelo_qwen3loop_sft_f16.gguf with huggingface_hub

Lordnyx
6b5ed511mo ago

Upload modelo_qwen3loop_sft_q8_0.gguf with huggingface_hub

Lordnyx
e5bdb561mo ago

Upload model.safetensors with huggingface_hub

Lordnyx
404b5bf1mo ago

Upload tokenizer_config.json with huggingface_hub

Lordnyx