Threads (@weststreetmorning)
local-llm
Original source

4B参数”无审查”模型推荐(基于Qwen系列abliteration改造)

Summary

Roundup of best uncensored 4B-parameter models, mostly based on Qwen series with “abliteration” (de-alignment) techniques.

Top picks:

  • 🏆 Qwen3-4B / Qwen3.5-4B (HauhauCS Aggressive) — best overall, aggressive de-censoring, 262k context, MMLU ~99.7%
  • ⚖️ Heretic variant (Qwen3-4B) — more refined technique, better preserves original capabilities
  • 🎭 JOSIE (Gabliterated) — full fine-tune on Qwen3-4B, personality-focused, natural conversation style for roleplay/companionship
  • 📱 Qwen3.5-4B-uncensored-MNN — mobile-optimized, MNN format 4-bit quant, ~2.5GB, 18-20 tok/s on flagship phones
  • 🧪 Gemma-3-4B-IT-Uncensored-v2 — Google Gemma 3 based, fast response, good instruction following

Other notable:

  • Huihui-Qwen3.5-4B-abliterated — popular community base model
  • Aisha-Qwen-Uncensored — Qwen3 in GGUF format for local inference
  • electroglyph/Qwen3-4B-Instruct-2507-uncensored

Selection tips:

  • Check benchmarks (Aggressive hits 99.7% MMLU)
  • Ensure format compatibility (GGUF/GPTQ) with your inference framework (Ollama/llama.cpp)
  • 4-bit quantization is good balance of performance vs size