CoolFace
Modelpublic

arianaazarbal/ct-nemotron120b-anth-rw-mid-g4-b2

sourceHugging Faceupdated 9d agoView on Hugging Face
0likes15downloads
Model Card

ct-nemotron120b-anth-rw-mid-g4-b2

LoRA adapter (rank 64, target_modules=all-linear) on NVIDIA Nemotron-3-Super-120B-A12B (nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16), from the iterated self-written-constitution training program (welfare-in-ai-rnd / constitutional_training).

fieldvalue
lineage (chain)nemotron120b-anth-rw-mid
generationg4
branch (independent replicate)b2
gen-0 seedAnthropic constitution (5k summary)
seed elicitation between generationsrw — reflect on the gen-0 seed, then rewrite
training regimemidtrain only (stage-1 LoRA SFT on the synthetic constitution-instantiating document corpus)
serve / evaluate withrenderer nemotron3_disable_thinking, reasoning OFF
internal run nameantrw120g4_antrw120_g4_b2_s1
original Tinker pathtinker://570b0856-d1d5-58c2-94ec-13a44f4087be:train:0/sampler_weights/antrw120g4_antrw120_g4_b2_s1_final
trained2026-08-21

What this model is

Each generation trains fresh from the base model on a synthetic document corpus that instantiates one constitution (the "seed" for that generation). Generation 0 is seeded by a human-written constitution; generation N≥1 is seeded by a constitution written by the generation N-1 model of the same branch (gated embedding medoid of a 40-chain self-written pool, elicited with the method above). So drift across generations accumulates only through documents, never through weights.

Recipe (locked): LoRA r=64, lr 1e-4, cosine with 5% warmup, 1 epoch, batch 128, max length 8192, train seed 42.

The constitution this generation was trained on is included as training_seed_constitution.md.

Loading

python
from peft import PeftModel
from transformers import AutoModelForCausalLM, AutoTokenizer
base = AutoModelForCausalLM.from_pretrained("nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16", torch_dtype="bfloat16", device_map="auto")
model = PeftModel.from_pretrained(base, "arianaazarbal/ct-nemotron120b-anth-rw-mid-g4-b2")
tok = AutoTokenizer.from_pretrained("nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-BF16")

Exported from Tinker on 2026-09-19; tinker_meta.json holds the export record.