renaudb1999/le-harnais-ft-agentworld-3b
060
le-harnais / ft-agentworld-3b
World-model student (balanced quality/speed).
- Base model: `meta-llama/Llama-3.2-3B-Instruct` — Built with Llama; Llama Community License applies.
- Class:
hero - Training data: datasets/agentworlddistilltrain.jsonl (160 ex; teacher=Qwen-AgentWorld-35B-A3B)
- Headline: token-F1 0.868 / OBSERVATION hit-rate 75% vs teacher (n=40)
Reproduce
BASE=meta-llama/Llama-3.2-3B-Instruct OUT=refs/llm-jepa/ft-agentworld-3b tools/distill_agentworld.shFull recipe, datasets, and eval commands: see docs/REPRODUCE.md in the [le-harnais distribution]. Provenance & license: docs/PROVENANCE.md.
Formats in this repo
*.safetensors— bf16 inference weights (serve withtransformersor le-harnaislh-serve/candle).*.Q4_K_M.gguf— portable 4-bit quant (run viaollama/llama.cpp; Mac-friendly).*.Q8_0.gguf— higher-fidelity 8-bit quant (hero models).
Orchestration amplifies a capable generator; it does not create competence.
