hattorihanzo1/cerberus-4b-gguf Q3_K_M GGUF - Free GGUF Download is indexed on GraySoft with repository links, GGUF quant files, and Hugging Face metadata. This page helps you pick a local model for guIDE or other runtimes. See related models in the same shard below.
Model Intelligence Sheet
hattorihanzo1/cerberus-4b-gguf overview
Cerberus-4B Cerberus non dormit — veritas sine pretio non datur...
Downloads
313
Likes
0
Pipeline
text-generation
Library
—
Visibility
Public
Access
Open
Repository Files & Downloads
9 files detected
Direct downloads for all repository files
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Cerberus-4b.F16.gguf | GGUF | F16 | 7.50 GB | Download |
| Cerberus-4b.IQ4_XS.gguf | GGUF | IQ4_XS | 2.13 GB | Download |
| Cerberus-4b.Q3_K_M.gguf | GGUF | Q3_K_M | 1.93 GB | Download |
| Cerberus-4b.Q4_K_M.gguf | GGUF | Q4_K_M | 2.33 GB | Download |
| Cerberus-4b.Q4_K_S.gguf | GGUF | Q4_K_S | 2.22 GB | Download |
| Cerberus-4b.Q5_K_M.gguf | GGUF | Q5_K_M | 2.69 GB | Download |
| Cerberus-4b.Q5_K_S.gguf | GGUF | Q5_K_S | 2.63 GB | Download |
| Cerberus-4b.Q6_K.gguf | GGUF | Q6_K | 3.08 GB | Download |
| Cerberus-4b.Q8_0.gguf | GGUF | — | 3.99 GB | Download |
Model Details Live
Metadata Inspector
Normalized metadata (stored in metadata_json)
{
"metadata": {},
"card_data": {
"language": [
"pl",
"en"
],
"license": "apache-2.0",
"base_model": "Qwen/Qwen3-4B",
"tags": [
"polish",
"chain-of-thought",
"reasoning",
"unsloth",
"llama-cpp",
"gguf",
"text-generation",
"qwen3",
"finetuned"
],
"pipeline_tag": "text-generation",
"frontmatter": {
"language": [
"pl",
"en"
],
"license": "apache-2.0",
"base_model": "Qwen/Qwen3-4B",
"tags": [
"polish",
"chain-of-thought",
"reasoning",
"unsloth",
"llama-cpp",
"gguf",
"text-generation",
"qwen3",
"finetuned"
],
"pipeline_tag": "text-generation"
},
"hero_image_url": "https://cdn-uploads.huggingface.co/production/uploads/68d1c6c3ea1c2d4e3c3df3f6/mMATq6mOBrzP5czbr5lSr.png",
"summary": "Cerberus-4B Cerberus non dormit — veritas sine pretio non datur...",
"quick_links": [],
"benchmark_table_html": "",
"readme_markdown": "---\nlanguage:\n - pl\n - en\nlicense: apache-2.0\nbase_model: Qwen/Qwen3-4B\ntags:\n - polish\n - chain-of-thought\n - reasoning\n - unsloth\n - llama-cpp\n - gguf\n - text-generation\n - qwen3\n - finetuned\npipeline_tag: text-generation\n---\n\n<p align=\"center\">\n <img src=\"https://cdn-uploads.huggingface.co/production/uploads/68d1c6c3ea1c2d4e3c3df3f6/mMATq6mOBrzP5czbr5lSr.png\" alt=\"Cerberus-4B\" width=\"800\"/>\n</p>\n<p align=\"center\">\n <em><strong><font size=\"8\" color=\"red\">Cerberus-4B</font></strong></em>\n <br>\n <a href=\"https://www.youtube.com/shorts/lTOMZEDjEC4\">\n <em><strong>Cerberus non dormit — veritas sine pretio non datur...</strong></em>\n </a>\n</p>\n\n## Κέρβερος — ὁ φύλαξ τῆς ἀληθείας\n*Cerber — strażnik prawdy*\nCerberus-4B to model językowy z natywnym wsparciem Chain-of-Thought, wykuty w treningu na starannie wyselekcjonowanych danych. Jak jego mityczny imiennik — nie przepuści byle czego. Każda odpowiedź poprzedzona jest rozumowaniem.\n\n## ⚔️ Geneza\nNikt nie rodzi się strażnikiem. Cerberus przeszedł przez wszystkie poziomy :\n\n| Faza | LR | Scheduler | Kroki | Cel |\n|------|-----|-----------|-------|-----|\n| I | 2e-4 | linear | 1500 | Wstępne opanowanie formatu CoT |\n| II | 3e-5 | constant | 1500 | Konsolidacja wiedzy i rozumowania |\n| III | 1e-5 | cosine | 1500 | Szlif — precyzja i głębia |\n| IV | humanistyczny | constant | 1500 | Dusza — język, finezja, polot |\n\nBaza: **Qwen3-4B** z natywnym tokenem `<think>` — architektura stworzona do rozumowania.\n\n## 🧠 Czym jest Cerberus?\n\n- **Polski model CoT** — myśli po polsku, rozumuje po polsku, odpowiada po polsku — ale dychy na piwo Ci nie pożyczy 🍺 \n- **Chain-of-Thought** — każda odpowiedź zawiera jawny proces myślowy w bloku `<think>`\n- **Wiedza ogólna + humanistyka** — nauki ścisłe, historia, filozofia, sztuka\n- **Wykształcony na destylowanych danych** — nie ilość, lecz jakość\n---\n## 💬 Format promptowania\n\n```\n<|im_start|>user\nTwoje pytanie tutaj<|im_end|>\n<|im_start|>assistant\n<think>\n...rozumowanie modelu...\n</think>\nOdpowiedź\n```\n## 📦 Dostępne kwantyzacje\n\n| Plik | Rozmiar | Zastosowanie |\n|------|---------|--------------|\n| Cerberus-4b.F16.gguf | ~8.0 GB | Referencyjna, pełna precyzja |\n| Cerberus-4b.Q8_0.gguf | ~4.3 GB | Wysoka jakość |\n| Cerberus-4b.Q6_K.gguf | ~3.3 GB | Zalecana — jakość vs rozmiar |\n| Cerberus-4b.Q5_K_M.gguf | ~2.9 GB | Dobry balans |\n| Cerberus-4b.Q5_K_S.gguf | ~2.7 GB | Szybsza wersja Q5 |\n| Cerberus-4b.Q4_K_M.gguf | ~2.5 GB | Codzienny użytek |\n| Cerberus-4b.Q4_K_S.gguf | ~2.4 GB | Lekka wersja Q4 |\n| Cerberus-4b.IQ4_XS.gguf | ~2.2 GB | Minimalistyczna |\n| Cerberus-4b.Q3_K_M.gguf | ~1.9 GB | Urządzenia mobilne |\n\n## 🔧 Uruchomienie (llama.cpp)\n```bash\nllama-cli \\\n -m Cerberus-4b.Q6_K.gguf \\\n -p \"<|im_start|>user\\nCzym jest absurd według Camusa?<|im_end|>\\n<|im_start|>assistant\\n\" \\\n -n 512 \\\n --temp 0.7 \\\n --repeat-penalty 1.1\n```\n## 🖥️ Wymagania sprzętowe\n| Kwantyzacja | Min. VRAM / RAM |\n|-------------|----------------|\n| Q4_K_M | 4 GB |\n| Q6_K | 6 GB |\n| Q8_0 | 8 GB |\n| F16 | 16 GB |\n\n## 📊 Dane treningowe\n\n- **Polski CoT** — wiedza ogólna, nauki ścisłe, lingwistyka, filozofia (25k rekordów)\n- **Polski instruct** — ogólny instruct po polsku (13k rekordów)\n- **Humanistyczny szlif** — sztuka, filozofia, finezja językowa (7k rekordów)\n\n## ⚠️ Ograniczenia\n- Model trenowany głównie na języku polskim — angielski działa ale nie jest priorytetem\n- Wiedza ograniczona do danych treningowych modelu bazowego (Qwen3-4B)\n- Nie zastępuje profesjonalnej porady medycznej, prawnej ani finansowej\n\n\n<p align=\"center\">\n <em><font size=\"8\"><strong> Τότε ἐν τῇ σκιᾷ μαχούμεθα </strong></font></em>\n <br>\n <em><font size=\"2\" color=\"silver\"><em></em>HattoriHanzo1 — Authentic Shinobi Tech ...</font></em>\n</p>\n",
"related_quantizations": []
},
"tags": [
"gguf",
"polish",
"chain-of-thought",
"reasoning",
"unsloth",
"llama-cpp",
"text-generation",
"qwen3",
"finetuned",
"pl",
"en",
"base_model:Qwen/Qwen3-4B",
"base_model:quantized:Qwen/Qwen3-4B",
"license:apache-2.0",
"endpoints_compatible",
"region:us",
"conversational"
],
"likes": 0,
"downloads": 313,
"gated": false,
"private": false,
"last_modified": "2026-03-22T16:45:17.000Z",
"created_at": "2026-03-22T14:36:53.000Z",
"pipeline_tag": "text-generation",
"library_name": ""
}
Source payload excerpt (from Hugging Face API)
{
"_id": "69bffe8586be13c318a0534c",
"id": "HattoriHanzo1/Cerberus-4B-GGUF",
"modelId": "HattoriHanzo1/Cerberus-4B-GGUF",
"sha": "e9c376863f1fd0dab1286175f6adadbd596cba38",
"createdAt": "2026-03-22T14:36:53.000Z",
"lastModified": "2026-03-22T16:45:17.000Z",
"author": "HattoriHanzo1",
"downloads": 313,
"likes": 0,
"gated": false,
"private": false,
"pipeline_tag": "text-generation",
"library_name": "",
"siblings_count": 11
}