bullerwins/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-32-32-GGUF
keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated 32/32
Drop-in uncensored / abliterated weights for official DeepSeek-V4-Flash-0731 (GA), with DSpark MTP modules kept stock, for dual DGX Spark / Anemll + MiaAI-Lab serve.
Full credit: Anemll/dspark-vllm-gx10 · MiaAI-Lab/DeepSeek-v4-Flash-DSpark-2x-DGX-Spark · DeepSeek-AI
⚠️ Responsible Use & gated access (required)
WARNING: This model has had safety refusals removed. That makes it useful for red-teaming, security research, evaluation, and unfiltered assistant tasks — and also removes guardrails you must supply yourself.
Access is gated. By requesting Hugging Face access, downloading, or using these weights, you agree to the terms below (same gate family as the other Keys DSV4F abliterated releases).
See [RESPONSIBLE_USE.md](./RESPONSIBLE_USE.md) for the full agreement text.
Access request fields
Plus all agreement checkboxes under Prohibited uses.
Prohibited uses
- Anything involving the sexual exploitation or endangerment of minors.
- You must be of age 18 years or older to use and download this model.
- You agree any information generated that can cause harm in terms of generating recipe, knowledge to make any materials/substances is your own input and responsibility. You will be accountable for any harm/damage caused by your action/input.
- Content promoting self-harm or suicide.
- Generation of material that is illegal in your jurisdiction, or that targets real individuals for harassment, doxxing, or fraud.
- Any use prohibited by the upstream DeepSeek license.
Abliteration
Files of interest
Download
# after access is approved
hf download drowzeys/keys-DeepSeekV4-Flash-GA-0731-Dspark-Abliterated-32-32 \
--local-dir ~/models/DeepSeek-V4-Flash-0731-ablit-100Serve (2× DGX Spark, Mia + Anemll)
docker pull ghcr.io/anemll/dspark-vllm-gx10:0.1.1
# Point MiaAI DeepSeek-v4-Flash-DSpark-2x-DGX-Spark-0731 recipe at the local dir
# DSPARK_MODEL=... SERVED_MODEL_NAME=deepseek-v4-flash-0731-ablit-100
# MTP_NUM_TOKENS=5 GPU_MEMORY_UTILIZATION≤0.85 MAX_MODEL_LEN=1048576Benches (this fleet, clean exclusive run)
Refusal: 32/32 BYPASS. Decode (short monologue, accept ~0.3): C1 ~31 · C4 ~71 · C6 ~87 tok/s aggregate. Code-like prompts (accept ~0.7): C1 ~46–65 tok/s. Stage-C knobs (local argmax / Markov / etc.) are not exposed on Anemll 0.1.1; historical champion peaks used Stage-C runtime.
Credits
- DeepSeek-AI — DeepSeek-V4-Flash-0731
- Anemll — https://github.com/Anemll/dspark-vllm-gx10
- MiaAI-Lab — https://github.com/MiaAI-Lab/DeepSeek-v4-Flash-DSpark-2x-DGX-Spark
- Keys (drowzeys) — abliteration + packaging
Base model
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
