HaniAI/SonaMath-0.5B
Add How to use (inference): download snippets, prompt format, greedy protocol
SonaMath-0.5B: rebrand from JunMath + SFT GSM8K full 13.8%
config: authoritative gsm8k_full_greedy 13.8% (182/1319)
Update eval: GSM8K full test 13.8% (182/1319), greedy, max_new=1024
Fix markdown: replace ~approx with 'approx.' (was rendered as strikethrough)
Config: release_stage=sft, rl_applied=false, split research tracks
Clarify SFT-only release vs separate RL research (address overclaim feedback)
Add rl_research metadata (GRPO alternative framing)
Clarify RL theory research as alternative to GRPO-style group optimization
Add theory_research field to public config
Add architecture-specific theory research section
Update config research_focus for SLM reasoning architecture
Emphasize new architecture research for small reasoning LMs
Release JunMath-0.5B research preview (weights + tokenizer + model card)
initial commit
