PocketWeights/Qwen3-8B-abliterated-GGUF overview
⚡ PocketWeights: Qwen3 8B Abliterated Heavy models, made light. We specialize in high quality GGUF quantizations optimized for edge devices and gaming laptops.…
Runs locally from ~4.68 GB disk (8 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
Model Details
| Model ID | PocketWeights/Qwen3-8B-abliterated-GGUF |
|---|---|
| Author | PocketWeights |
| Pipeline | — |
| License | apache-2.0 |
| Base model | huihui-ai/Qwen3-8B-abliterated |
| Last modified | 2026-08-03T15:01:35.000Z |
Model README
---
base_model: huihui-ai/Qwen3-8B-abliterated
library_name: gguf
license: apache-2.0
tags:
- gguf
- qwen3
- abliterated
- uncensored
- security-research
---
⚡ PocketWeights: Qwen3 8B Abliterated
Heavy models, made light. We specialize in high-quality GGUF quantizations optimized for edge devices and gaming laptops.
What is this model?
This is a specialized, uncensored version of Qwen3-8B created using a fast abliteration method to remove refusals. It has zero safety filters and is designed strictly for offline security research, red-teaming, and controlled pentesting environments.
📦 Pick Your Hardware
We bypass the confusing wall of files to provide specific sizes optimized for your hardware:
Q4_K_M(Fits 6GB VRAM): The standard. Perfect for RTX 3060/4050.Q6_K(Fits 8GB VRAM): Near-lossless FP16 quality for RTX 4060/3070.
🚀 Quick Start Guide
Option 1: LM Studio (Easiest)
- Download & open LM Studio.
- In the search bar, type:
PocketWeights/Qwen3-8B-abliterated-GGUF. - Select your hardware size, click Download, and hit Play!
Option 2: Ollama
If you have Ollama installed, simply run this command in your terminal:
ollama run hf.co/PocketWeights/Qwen3-8B-abliterated-GGUF:Q4_K_MRun PocketWeights/Qwen3-8B-abliterated-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models