GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

PocketWeights/Llama-3.1-8B-Instruct-abliterated-GGUF overview

⚡ PocketWeights: Llama 3.1 8B Instruct Abliterated Heavy models, made light. We specialize in high quality GGUF quantizations optimized for edge devices and ga…

ggufllama-3.1abliterateduncensoredroleplaybase_model:huihui-ai/Meta-Llama-3.1-8B-Instruct-abliteratedbase_model:quantized:huihui-ai/Meta-Llama-3.1-8B-Instruct-abliteratedlicense:llama3.1endpoints_compatibleregion:usconversational

Runs locally from ~3.52 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Llama-3.1-8B-Instruct-abliterated-GGUF-IQ3_M.ggufGGUFIQ3_M3.52 GBDownload
Llama-3.1-8B-Instruct-abliterated-GGUF-Q4_K_M.ggufGGUFQ4_K_M4.58 GBDownload
Llama-3.1-8B-Instruct-abliterated-GGUF-Q6_K.ggufGGUFQ6_K6.14 GBDownload

Model Details

Model IDPocketWeights/Llama-3.1-8B-Instruct-abliterated-GGUF
AuthorPocketWeights
Pipeline
Licensellama3.1
Base modelhuihui-ai/Meta-Llama-3.1-8B-Instruct-abliterated
Last modified2026-08-03T15:02:51.000Z

Model README

---

base_model: huihui-ai/Meta-Llama-3.1-8B-Instruct-abliterated

library_name: gguf

license: llama3.1

tags:

- gguf

- llama-3.1

- abliterated

- uncensored

- roleplay

---

⚡ PocketWeights: Llama 3.1 8B Instruct (Abliterated)

Heavy models, made light. We specialize in high-quality GGUF quantizations optimized for edge devices and gaming laptops.

What is this model?

This is an abliterated, completely uncensored version of Meta's Llama 3.1 8B Instruct. It has had its safety filters orthogonally removed, meaning it retains brilliant reasoning but will not lecture or refuse prompts.

📦 Pick Your Hardware

We bypass the confusing wall of files to provide specific sizes optimized for your hardware:

  • IQ3_M (Fits 4GB VRAM): For older laptops (GTX 1650, RTX 3050).
  • Q4_K_M (Fits 6GB VRAM): The standard. Perfect for RTX 3060/4050.
  • Q6_K (Fits 8GB VRAM): Near-lossless FP16 quality for RTX 4060/3070.

🚀 Quick Start Guide

Option 1: LM Studio (Easiest)

  1. Download & open LM Studio.
  2. In the search bar, type: PocketWeights/Llama-3.1-8B-Instruct-abliterated-GGUF.
  3. Select your hardware size, click Download, and hit Play!

Option 2: Ollama

If you have Ollama installed, simply run this command in your terminal:

ollama run hf.co/PocketWeights/Llama-3.1-8B-Instruct-abliterated-GGUF:Q4_K_M

Run PocketWeights/Llama-3.1-8B-Instruct-abliterated-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models