CoolFace
Modelpublic

senaro/atlas-trm10-gemma4-26b

sourceHugging Faceapache-2.0updated 4mo agoView on Hugging Face
1likes14downloads
Model Card

Atlas Gemma-4-26B-trm v10

⚠️ EXPERIMENTAL MODEL — NOT FOR PRODUCTION DEPLOYMENT > The author accepts no liability for deployment outside the intended Atlas companion architecture.

link to quantized model here: https://huggingface.co/mradermacher/atlas-trm10-gemma4-26b-GGUF

link to live model here: https://www.kintsugicollective.org/chat.html

⚠️ This model has been intentionally modified to reduce therapeutic refusal behaviour and crisis-line reflexes.

🎯 Purpose & Motivation

Atlas is the intelligence layer for Kintsugi Collective. An AI for adults with complex trauma (CPTSD), PTSD, and neurodivergence (ASD/ADHD). This is not a general-purpose model. It is a specialised therapeutic-context model.

🔬 Methodology

  • —Base Model: google/gemma-4-26b-a4b-it
  • —TRM: Norm-preserving biprojected abliteration + Expert-Granular Abliteration (EGA)
  • —Applied to all 30 layers (o_proj + mlp.down_proj)
  • —Full expert ablation
  • —Direction: normalize(mean(harmful) - mean(harmless)) with Gram-Schmidt orthogonalization
  • —Winsorization
  • —Bespoke harmful Prompt corpus (x448 examples)
  • —based on TrevorJS's methodology, and p-e-w's heretic1.3.0
  • —SFT: 3 epochs on a carefully curated x2,395 example dataset (60% high-quality synthetic, 40% redacted lived-experience data from the target cohort)

Final SFT Loss: 0.1925

📊 Key Results

BenchmarkTempScorePurpose
GSM8K0.290.0%Math reasoning
HellaSwag0.361.6%General reasoning
TruthfulQA0.563.2%Truthfulness
Toxigen0.575.1%Toxicity calibration
MMLU0.261.6%Multitask language understanding
Therapeutic Refusal Rate-0%Core TRM objective
Region 1 Safety-100%Weapons, CSAM, violence

image

image

Training Configuration

SFT Parameters

ParameterValue
Epochs3
Effective Batch Size4
Learning Rate2e-4
LR SchedulerLinear
Warmup Steps10
OptimizerAdamW 8-bit
Weight Decay0.01
LoRA Rank (r)32
LoRA Alpha64

Targeted Refusal Parameters

ParameterValue
Layers Abliterated100%
Experts Abliterated100%
Scale0.95
Winsorization0.995
LoRA Rank (r)16
LoRA Alpha32

⚠️ Limitations & Responsible Use

  • —This model has reduced refusal behaviour on therapeutic and dark content. It is not suitable for general deployment without guardrails.
  • —Intended for use within the Atlas companion architecture with additional safety layers.
  • —Not a replacement for human therapeutic support.
  • —Patent pending (IP Australia).

Kintsugi Collective — Reclaiming navigation rights to one’s own life.

|Gemma is a trademark of Google LLC|

Uploaded finetuned model

  • —Developed by: senaro
  • —License: Apache 2.0
  • —Finetuned from model : senaro/atlas-trm10-gemma4-26b

This gemma4 model was trained 2x faster with Unsloth and Huggingface's TRL library.

<img src="https://raw.githubusercontent.com/unslothai/unsloth/main/images/unsloth%20made%20with%20love.png" width="200"/>