ApiFort/LLMFort-jailbreak_content_injection
π‘οΈ LLM-Fort Guardrails Suite (v1)
  
LLM-Fort Guardrails is a suite of 7 security-focused LoRA adapters fine-tuned on top of Qwen/Qwen3-4B-Instruct-2507. These adapters serve as lightweight, high-performance security guardrails mapped to critical safety boundaries.
By offloading classification and security checks to lightweight adapters, the system achieves enterprise-grade security filtering without degrading the inference performance of the main application model.
π Collection Page: ApiFort/llmfort-guardrails-v1
πΊοΈ Category Mappings
π Performance & Evaluation
Visual comparison of baseline performance versus the trained adapters:
Benchmark Results
Below is the exact accuracy performance measured across our evaluation test suites:
ποΈ Training & Validation Datasets
The adapters were trained and validated on the following dataset references:
