GraySoft
Projects Models About FAQ Contact Download guIDE →

hattorihanzo1/cerberus-4b-gguf 4b.Q8_0 GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.

Model Intelligence Sheet

hattorihanzo1/cerberus-4b-gguf overview

Cerberus-4B Cerberus non dormit — veritas sine pretio non datur...

ggufpolishchain-of-thoughtreasoningunslothllama-cpptext-generationqwen3finetunedplenbase_model:Qwen/Qwen3-4Bbase_model:quantized:Qwen/Qwen3-4Blicense:apache-2.0endpoints_compatibleregion:usconversational
hattorihanzo1/cerberus-4b-gguf visual
Downloads
313
Likes
0
Pipeline
text-generation
Library
Visibility
Public
Access
Open

Repository Files & Downloads

9 files detected
Direct downloads for all repository files
FileTypeQuantizationSizeLink
Cerberus-4b.F16.gguf GGUF F16 7.50 GB Download
Cerberus-4b.IQ4_XS.gguf GGUF IQ4_XS 2.13 GB Download
Cerberus-4b.Q3_K_M.gguf GGUF Q3_K_M 1.93 GB Download
Cerberus-4b.Q4_K_M.gguf GGUF Q4_K_M 2.33 GB Download
Cerberus-4b.Q4_K_S.gguf GGUF Q4_K_S 2.22 GB Download
Cerberus-4b.Q5_K_M.gguf GGUF Q5_K_M 2.69 GB Download
Cerberus-4b.Q5_K_S.gguf GGUF Q5_K_S 2.63 GB Download
Cerberus-4b.Q6_K.gguf GGUF Q6_K 3.08 GB Download
Cerberus-4b.Q8_0.gguf GGUF 3.99 GB Download

Model Details Live

Model Slug
hattorihanzo1/cerberus-4b-gguf
Author
HattoriHanzo1
Pipeline Task
text-generation
Library
Created
2026-03-22
Last Modified
2026-03-22
Gated
No
Private
No
HF SHA
e9c376863f1fd0dab1286175f6adadbd596cba38
License
apache-2.0
Language
pl, en
Base Model
Qwen/Qwen3-4B

Metadata Inspector

Normalized metadata (stored in metadata_json)
{
  "metadata": {},
  "card_data": {
    "language": [
      "pl",
      "en"
    ],
    "license": "apache-2.0",
    "base_model": "Qwen/Qwen3-4B",
    "tags": [
      "polish",
      "chain-of-thought",
      "reasoning",
      "unsloth",
      "llama-cpp",
      "gguf",
      "text-generation",
      "qwen3",
      "finetuned"
    ],
    "pipeline_tag": "text-generation",
    "frontmatter": {
      "language": [
        "pl",
        "en"
      ],
      "license": "apache-2.0",
      "base_model": "Qwen/Qwen3-4B",
      "tags": [
        "polish",
        "chain-of-thought",
        "reasoning",
        "unsloth",
        "llama-cpp",
        "gguf",
        "text-generation",
        "qwen3",
        "finetuned"
      ],
      "pipeline_tag": "text-generation"
    },
    "hero_image_url": "https://cdn-uploads.huggingface.co/production/uploads/68d1c6c3ea1c2d4e3c3df3f6/mMATq6mOBrzP5czbr5lSr.png",
    "summary": "Cerberus-4B   Cerberus non dormit — veritas sine pretio non datur...",
    "quick_links": [],
    "benchmark_table_html": "",
    "readme_markdown": "---\nlanguage:\n  - pl\n  - en\nlicense: apache-2.0\nbase_model: Qwen/Qwen3-4B\ntags:\n  - polish\n  - chain-of-thought\n  - reasoning\n  - unsloth\n  - llama-cpp\n  - gguf\n  - text-generation\n  - qwen3\n  - finetuned\npipeline_tag: text-generation\n---\n\n<p align=\"center\">\n  <img src=\"https://cdn-uploads.huggingface.co/production/uploads/68d1c6c3ea1c2d4e3c3df3f6/mMATq6mOBrzP5czbr5lSr.png\" alt=\"Cerberus-4B\" width=\"800\"/>\n</p>\n<p align=\"center\">\n  <em><strong><font size=\"8\" color=\"red\">Cerberus-4B</font></strong></em>\n  <br>\n  <a href=\"https://www.youtube.com/shorts/lTOMZEDjEC4\">\n    <em><strong>Cerberus non dormit — veritas sine pretio non datur...</strong></em>\n  </a>\n</p>\n\n## Κέρβερος — ὁ φύλαξ τῆς ἀληθείας\n*Cerber — strażnik prawdy*\nCerberus-4B to model językowy z natywnym wsparciem Chain-of-Thought, wykuty w treningu na starannie wyselekcjonowanych danych. Jak jego mityczny imiennik — nie przepuści byle czego. Każda odpowiedź poprzedzona jest rozumowaniem.\n\n## ⚔️ Geneza\nNikt nie rodzi się strażnikiem. Cerberus przeszedł przez wszystkie poziomy :\n\n| Faza | LR | Scheduler | Kroki | Cel |\n|------|-----|-----------|-------|-----|\n| I | 2e-4 | linear | 1500 | Wstępne opanowanie formatu CoT |\n| II | 3e-5 | constant | 1500 | Konsolidacja wiedzy i rozumowania |\n| III | 1e-5 | cosine | 1500 | Szlif — precyzja i głębia |\n| IV | humanistyczny | constant | 1500 | Dusza — język, finezja, polot |\n\nBaza: **Qwen3-4B** z natywnym tokenem `<think>` — architektura stworzona do rozumowania.\n\n## 🧠 Czym jest Cerberus?\n\n- **Polski model CoT** — myśli po polsku, rozumuje po polsku, odpowiada po polsku — ale dychy na piwo Ci nie pożyczy 🍺 \n- **Chain-of-Thought** — każda odpowiedź zawiera jawny proces myślowy w bloku `<think>`\n- **Wiedza ogólna + humanistyka** — nauki ścisłe, historia, filozofia, sztuka\n- **Wykształcony na destylowanych danych** — nie ilość, lecz jakość\n---\n## 💬 Format promptowania\n\n```\n<|im_start|>user\nTwoje pytanie tutaj<|im_end|>\n<|im_start|>assistant\n<think>\n...rozumowanie modelu...\n</think>\nOdpowiedź\n```\n## 📦 Dostępne kwantyzacje\n\n| Plik | Rozmiar | Zastosowanie |\n|------|---------|--------------|\n| Cerberus-4b.F16.gguf | ~8.0 GB | Referencyjna, pełna precyzja |\n| Cerberus-4b.Q8_0.gguf | ~4.3 GB | Wysoka jakość |\n| Cerberus-4b.Q6_K.gguf | ~3.3 GB | Zalecana — jakość vs rozmiar |\n| Cerberus-4b.Q5_K_M.gguf | ~2.9 GB | Dobry balans |\n| Cerberus-4b.Q5_K_S.gguf | ~2.7 GB | Szybsza wersja Q5 |\n| Cerberus-4b.Q4_K_M.gguf | ~2.5 GB | Codzienny użytek |\n| Cerberus-4b.Q4_K_S.gguf | ~2.4 GB | Lekka wersja Q4 |\n| Cerberus-4b.IQ4_XS.gguf | ~2.2 GB | Minimalistyczna |\n| Cerberus-4b.Q3_K_M.gguf | ~1.9 GB | Urządzenia mobilne |\n\n## 🔧 Uruchomienie (llama.cpp)\n```bash\nllama-cli \\\n  -m Cerberus-4b.Q6_K.gguf \\\n  -p \"<|im_start|>user\\nCzym jest absurd według Camusa?<|im_end|>\\n<|im_start|>assistant\\n\" \\\n  -n 512 \\\n  --temp 0.7 \\\n  --repeat-penalty 1.1\n```\n## 🖥️ Wymagania sprzętowe\n| Kwantyzacja | Min. VRAM / RAM |\n|-------------|----------------|\n| Q4_K_M | 4 GB |\n| Q6_K | 6 GB |\n| Q8_0 | 8 GB |\n| F16 | 16 GB |\n\n## 📊 Dane treningowe\n\n- **Polski CoT** — wiedza ogólna, nauki ścisłe, lingwistyka, filozofia (25k rekordów)\n- **Polski instruct** — ogólny instruct po polsku (13k rekordów)\n- **Humanistyczny szlif** — sztuka, filozofia, finezja językowa (7k rekordów)\n\n## ⚠️ Ograniczenia\n- Model trenowany głównie na języku polskim — angielski działa ale nie jest priorytetem\n- Wiedza ograniczona do danych treningowych modelu bazowego (Qwen3-4B)\n- Nie zastępuje profesjonalnej porady medycznej, prawnej ani finansowej\n\n\n<p align=\"center\">\n  <em><font size=\"8\"><strong> Τότε ἐν τῇ σκιᾷ μαχούμεθα </strong></font></em>\n  <br>\n  <em><font size=\"2\" color=\"silver\"><em></em>HattoriHanzo1 — Authentic Shinobi Tech ...</font></em>\n</p>\n",
    "related_quantizations": []
  },
  "tags": [
    "gguf",
    "polish",
    "chain-of-thought",
    "reasoning",
    "unsloth",
    "llama-cpp",
    "text-generation",
    "qwen3",
    "finetuned",
    "pl",
    "en",
    "base_model:Qwen/Qwen3-4B",
    "base_model:quantized:Qwen/Qwen3-4B",
    "license:apache-2.0",
    "endpoints_compatible",
    "region:us",
    "conversational"
  ],
  "likes": 0,
  "downloads": 313,
  "gated": false,
  "private": false,
  "last_modified": "2026-03-22T16:45:17.000Z",
  "created_at": "2026-03-22T14:36:53.000Z",
  "pipeline_tag": "text-generation",
  "library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
  "_id": "69bffe8586be13c318a0534c",
  "id": "HattoriHanzo1/Cerberus-4B-GGUF",
  "modelId": "HattoriHanzo1/Cerberus-4B-GGUF",
  "sha": "e9c376863f1fd0dab1286175f6adadbd596cba38",
  "createdAt": "2026-03-22T14:36:53.000Z",
  "lastModified": "2026-03-22T16:45:17.000Z",
  "author": "HattoriHanzo1",
  "downloads": 313,
  "likes": 0,
  "gated": false,
  "private": false,
  "pipeline_tag": "text-generation",
  "library_name": "",
  "siblings_count": 11
}