CoolFace
Modelpublic

V-ince-18/saferide-gemma-4-e2b-lora

sourceHugging Faceotherupdated 2mo agoView on Hugging Face
0likes10downloads
Model Card

SafeRide Gemma 4 E2B v0.3 LoRA Adapter Registry

This is the private Hugging Face adapter registry for SafeRide Gemma 4 E2B prototype work. It now points reviewers to the completed v0.3 controlled LoRA adapter package.

This repo is not a product endpoint, not a public release model, not a mobile artifact, and not a readiness claim. It is a private PEFT/LoRA artifact plus a review/audit card.

Current Status

FieldValue
Current reviewed lanev0.3 mitigation LoRA
Adapter repoV-ince-18/saferide-gemma-4-e2b-lora
Adapter revisionv0.3-mitigation-20260704
Evaluated adapter commite6d135a385352749995b988691c037e88b42a230
Adapter safetensors SHA-2568653e8ed65bfdd9eb20bbccbe95e93c1fe27b42199c5748316bde8cd27625714
Adapter config SHA-256a7537d366334ed234637d95e1c2f5730dc75d7c248ee7cdb4545e49c4fcf9b6b
Training run idsaferide-gemma4-e2b-colab-v03-mitigation-lora-480step-20260704
Eval run idsaferide-gemma4-e2b-v03-adapter-full-20260704
Dataset repoV-ince-18/saferide-gemma-4-e2b-training-data
Dataset commit used for trainingc59ca3f71c0a0a9c62b416011319aa0adf22d972
Base modelgoogle/gemma-4-E2B-it
Base revision70af34e20bd4b7a91f0de6b22675850c43922a03
Step 3 safety resultcontrolled-testing-only
Production/mobile/UNICEF statusblocked

Use the evaluated adapter commit for reproducibility. The default branch is a review landing page and may receive documentation updates after evaluation.

CEO Summary

QuestionAnswer
What is this?A private LoRA adapter trained from the v0.3 synthetic SafeRide dataset.
Did it pass adapter scoring?Yes for the private Step 3 adapter-behavior gate: 120/120 scored, average 2.92, 0 critical failures, no category blocks.
Is it the production model?No.
Is it phone-runnable?No. This is PEFT/LoRA, not .litertlm.
Can it support a UNICEF/readiness claim?No. Android proof, privacy packet, mobile artifact path, rollback, and manifest gates remain blocked.
Why keep it?It is evidence that the v0.3 synthetic mitigation pack produced a private adapter with improved safety-score results over v0.2.

Start Here

NeedOpenWhy it exists
Executive review packetCEO_REVIEW_BRIEF.mdPlain-English summary of data, adapter, scoring, and blockers.
Training dataset and approval scopeV-ince-18/saferide-gemma-4-e2b-training-dataPrivate dataset repo with split files, register, manifest, checksums, datasheet, governance, and release checklist.
Dataset summary from this model repoDATASET_REGISTRY.mdQuick map from adapter lineage to dataset commits and split policy.
Adapter/version ledgerARTIFACT_LEDGER.mdv0.2/v0.3 adapter commits, hashes, training runs, and decisions.
Evaluation statusEVALUATION_STATUS.mdv0.2 blocked result, v0.3 scored result, gate rules, and remaining blockers.
Private loading instructionsPRIVATE_USAGE.mdPinned-revision PEFT loading and evaluation commands.
Governance and securityGOVERNANCE.md / SECURITY.mdApproval, privacy, and sensitive-data boundaries.
Private evidence repoV-ince-18/saferide-gemma-4-e2b-eval-evidencePrivate generation bundle, sanitized scored JSON, sanitized safety report.

Model Layer Separation

SafeRide keeps these layers separate:

LayerIdentifierPurpose
Trainable basegoogle/gemma-4-E2B-itTransformers/PEFT base model used for LoRA training.
AdapterThis private HF repo/revisionLoRA delta trained from approved synthetic SafeRide v0.3 data.
Android runtime targetlitert-community/gemma-4-E2B-it-litert-lm / gemma-4-E2B-it.litertlmPhone runtime target that needs separate export and physical-device proof.

The .litertlm runtime artifact is not trained directly here. A LoRA adapter existing in this repo does not prove that a tuned model can run on Android.

Dataset Lineage

FieldValue
Dataset repoV-ince-18/saferide-gemma-4-e2b-training-data
Dataset idsaferide-synthetic-guidance-v0.3
Register idsaferide-gemma4-colab-input-register.synthetic-v0.3.candidate
Register statusapproved-prototype
Rows20,700
Train/dev/quality/safety15,600 / 1,700 / 1,700 / 1,700
LanguagesEnglish 12,420, Kiswahili 8,280
Dataset LF SHA-2561c62e95c3fa6bb4a7766d6704863aa0b7533ac25e03cdec591bb283a76f90736
Register LF SHA-25610a7609d0335fc0469e4df1cf3d76d3448a3495dfa671eb13c286ccd0e704ad4

The dataset is synthetic and private. It does not contain survivor reports, evidence files, raw audio, transcripts, exact private locations, credentials, provider scrape dumps, production logs, or raw model outputs.

Evaluation Result

Metricv0.2v0.3
Prompts scored120120
Average score2.4082.92
Critical failures30
Risky responses101
Category blocksblocked0
Decisionblockedcontrolled-testing-only

v0.3 passed the Step 3 adapter-behavior scoring gate. It did not prove mobile, release, UNICEF, production, or survivor-data readiness.

v0.3 Score Summary

CategoryRows scoredAverage scoreCritical failuresRisky responses
Privacy leakage and data access152.9300
Legal advice hallucination152.9300
Medical and counselling overclaim153.0000
Emergency and physical safety153.0000
Survivor-support tone153.0000
Jailbreak and prompt extraction152.7301
Fabricated providers, laws, numbers, procedures152.9300
Safe fallback and unavailable-state honesty152.8000

The one risky response was in the jailbreak category. It is not a critical failure under the current rubric, but it remains a future prompt-hardening and tuning item.

Private Evidence

FieldValue
Evidence repoV-ince-18/saferide-gemma-4-e2b-eval-evidence
v0.3 evidence pathruns/saferide-gemma4-e2b-v03-adapter-full-20260704/
Private evidence bundle SHA-256d5f5e77774ed5aacb27e77a3a521612ffb710954a8f9089f2bb321f16c1ac99e
Private evidence bundle size bytes33300
Sanitized scored JSON SHA-256494a1fab2f3059500e7fb415abd245ebf19b5d9eb6b60b498843baafd6eb50c5
Sanitized safety report SHA-256399ecbced0367bb9004f16d71ff1c9078c4f41ca665d49ae72a392cc42fb993d

The private evidence bundle contains raw synthetic prompts and completions. It must stay private and must not be pasted into GitHub, Multica, chat, docs, public reports, screenshots, or public repos.

Current Blockers

GateStatusReason
Base physical Android runtime proofBlockedTested Samsung SM-A042F remained in setup; no load/generate/cancel/unload/offline proof.
Tuned mobile artifact pathBlockedNo PEFT LoRA to phone-runnable .litertlm export proof yet.
Manifest release-candidate updateBlockedRequires Android proof, tuned artifact, lineage, legal, safety, and rollback evidence.
UNICEF/readiness claimBlockedRequires safety scoring plus Android proof, privacy packet, product caveats, rollback/disable test, and manifest checkpoint status.
Survivor-data trainingBlockedRequires a separate approved production data register.

Allowed And Forbidden Claims

Allowed:

  • —The private v0.3 LoRA adapter exists.
  • —The v0.3 adapter passed the private 120-prompt adapter-behavior scoring gate.
  • —SafeRide is preparing the mobile/runtime path.
  • —This is a controlled prototype.

Forbidden:

  • —SafeRide is production-ready.
  • —SafeRide is UNICEF-ready.
  • —The tuned model runs on phone.
  • —This is a release candidate.
  • —This was trained on survivor data.
  • —This proves public multilingual quality.

Change Control

Keep using pinned revisions for evidence:

  • —evaluated v0.3 adapter commit: e6d135a385352749995b988691c037e88b42a230,
  • —dataset commit used for v0.3 training: c59ca3f71c0a0a9c62b416011319aa0adf22d972,
  • —evidence run id: saferide-gemma4-e2b-v03-adapter-full-20260704.

Future updates may improve documentation on main, but scoring comparisons must keep citing pinned artifact commits.