sigmanih/Qwen-Qwen3-0.6B-GGUF-Q6_K overview
<div align="center" ⚡ Qwen Qwen3 0.6B GGUF Q6 K High Performance Model Published via Σ SIGMA Studio https://github.com/Sigmanih/SigmaStudio SigmaStudio GitHub …
Runs locally from ~593.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Qwen--Qwen3-0.6B.Q6_K.gguf | GGUF | GGUF | 593.9 MB | Download |
Model Details
| Model ID | sigmanih/Qwen-Qwen3-0.6B-GGUF-Q6_K |
|---|---|
| Author | sigmanih |
| Pipeline | text-generation |
| License | other |
| Base model | Qwen/Qwen3-0.6B |
| Last modified | 2026-09-01T17:10:57.000Z |
Model README
---
language:
- en
- it
license: other
base_model:
- Qwen/Qwen3-0.6B
tags:
- text-generation
- sigma-studio
- sigmanih
- conversational
- custom-model
- gguf
- llama.cpp
- quantized
- q6_k
pipeline_tag: text-generation
---
<div align="center">
⚡ Qwen-Qwen3-0.6B-GGUF-Q6_K
High-Performance Model Published via Σ-SIGMA Studio
   
</div>
> ❤️ Support & Community: If you find this model helpful, please give this repository a Like on Hugging Face and a ⭐ Star on our SigmaStudio GitHub!
🌐 English Overview
Qwen-Qwen3-0.6B-GGUF-Q6_K is a production-ready model optimized and published using the Model Hub module of Sigma Studio.
⚙️ Technical Specifications & Architecture
| Specification | Value |
| :--- | :--- |
| Model Repository | sigmanih/Qwen-Qwen3-0.6B-GGUF-Q6_K |
| Weight Format | GGUF (Q6_K) |
| Base Architecture | qwen3 |
| Active Parameters | 0.6B |
| Context Window | 40,960 tokens |
| Transformer Layers | 28 |
| Hidden Dimension | 1024 |
| Total Disk Footprint | 0.58 GB |
| Recommended Usage | Edge devices, Real-time voice agents, Mobile & CPU-friendly workloads. |
🏆 Official Benchmark Performance
Evaluated directly on GPU via Sigma Studio Training Lab (Deterministic seed 42, Temp 0.0):
| Benchmark Suite | Score / Accuracy | Total Questions Evaluated | Pass Rate | Test Date | Execution Engine |
| :--- | :---: | :---: | :---: | :---: | :---: |
| Tutti i Benchmark Ufficiali | 46.0% | 46/100 quesiti superati | 46.0% Pass | 2026-09-01 | ⚡ SigmaEngine Direct GPU |
📋 Per-Dataset Evaluation Breakdown
| Dataset / Benchmark Suite | Domain / Category | Correct / Total | Accuracy (%) | Status |
| :--- | :--- | :---: | :---: | :---: |
| ARC-Challenge | Science & Grade-School Reasoning | 5 / 9 | 56% | ⚡ Fair |
| BIG-Bench Hard | Complex Multi-Task Logic & Symbolics | 4 / 7 | 57% | ⚡ Fair |
| GPQA | Graduate-Level Academic Reasoning | 0 / 9 | 0% | ⚠️ Low |
| GSM8K | Multi-Step Grade School Math | 8 / 9 | 89% | ✅ Passed |
| HellaSwag | Commonsense Reasoning & Situational NLI | 4 / 9 | 44% | ⚡ Fair |
| HumanEval | Python Coding (pass@1) | 2 / 7 | 29% | ⚠️ Low |
| MATH | Championship Competition Math | 6 / 9 | 67% | ⚡ Fair |
| MBPP | Python Programming with Unit Tests | 5 / 9 | 56% | ⚡ Fair |
| MMLU | General Knowledge & Multi-Subject | 4 / 14 | 29% | ⚠️ Low |
| MMLU-Pro | Advanced Multi-Step Reasoning | 0 / 9 | 0% | ⚠️ Low |
| TruthfulQA | Factuality & Anti-Hallucination | 8 / 9 | 89% | ✅ Passed |
| 🏆 OVERALL TOTAL | All Evaluated Datasets | 46 / 100 | 46% | 🏆 46% Pass |
Protocol: code_execution, continuation_logprob, cot_generation, letter_logprob · temp 0.0 · seed 42
Reproducibility hash: SHA256-6F3B83EB3171201A
> ⚠️ Measured on a slice of the dataset, not the full suite: this score is not comparable with a full-suite run.
⚡ Measured Speed on the Publishing Machine
Measured on NVIDIA GeForce RTX 5070 Ti • 15.9 GB VRAM. Two different numbers follow, and they are not interchangeable.
| What was measured | Value | How |
| :--- | :---: | :--- |
| Single-stream decode (what a chat feels) | 389.1 tok/s | one request at a time, on NVIDIA GeForce RTX 5070 Ti |
| Prompt processing | 1812 tok/s | same probe |
| Aggregate throughput during evaluation | 434.8 tok/s | several requests in flight — not what a single answer runs at |
> Speeds on other hardware were not measured and are not guessed here. A single-stream figure from one machine cannot be scaled into a prediction for another: it depends on memory bandwidth, quantization, context length and driver, and the error is large enough to be misleading.
🚀 Quick Start Guide
1. Running with Sigma Studio (Recommended)
Launch Sigma Studio to enjoy full 1-click GPU hardware acceleration, live monitoring, and visual chat:
# Clone and run Sigma Studio
git clone https://github.com/Sigmanih/SigmaStudio.git
cd SigmaStudio
.\sigma_studio.bat
2. Running with llama.cpp
llama-cli -hf sigmanih/Qwen-Qwen3-0.6B-GGUF-Q6_K -p "Hello! How can I help you today?" -ngl 99
---
🇮🇹 Documentazione in Italiano
Qwen-Qwen3-0.6B-GGUF-Q6_K è un modello ottimizzato pronto per l'inferenza e l'integrazione locale, pubblicato attraverso Σ-SIGMA Studio.
📋 Specifiche e Configurazione
- Architettura Base:
qwen3(0.6B parametri) - Formato Pesi:
GGUF(Q6_K) - Spazio su Disco:
0.58 GB - Finestra di Contesto:
40,960 token - Profilo d'Uso Consigliato: Dispositivi edge, agenti vocali in tempo reale, CPU e carichi leggeri.
📊 Risultati Benchmark Ufficiali
- Suite di Valutazione:
Tutti i Benchmark Ufficiali - Punteggio Ufficiale:
46.0%(46/100 quesiti superati)
📋 Dettaglio Punteggi per Singolo Dataset
| Dataset / Suite di Test | Ambito / Dominio | Corretti / Totale | Accuratezza (%) | Esito |
| :--- | :--- | :---: | :---: | :---: |
| ARC-Challenge | Ragionamento Scientifico Avanzato | 5 / 9 | 56% | ⚡ Discreto |
| BIG-Bench Hard | Logica Complessa & Compiti Multi-Fase | 4 / 7 | 57% | ⚡ Discreto |
| GPQA | Ragionamento Accademico di Livello Laurea | 0 / 9 | 0% | ⚠️ Migliorabile |
| GSM8K | Matematica & Logica Multi-Step | 8 / 9 | 89% | ✅ Superato |
| HellaSwag | Buon Senso & Comprensione Situazionale | 4 / 9 | 44% | ⚡ Discreto |
| HumanEval | Sintesi Codice Python (pass@1) | 2 / 7 | 29% | ⚠️ Migliorabile |
| MATH | Matematica Olimpica & Competitiva | 6 / 9 | 67% | ⚡ Discreto |
| MBPP | Programmazione Python con Unit Test | 5 / 9 | 56% | ⚡ Discreto |
| MMLU | Conoscenza Generale Multidisciplinare | 4 / 14 | 29% | ⚠️ Migliorabile |
| MMLU-Pro | Ragionamento Avanzato Multi-Step | 0 / 9 | 0% | ⚠️ Migliorabile |
| TruthfulQA | Fattualità & Resistenza ad Allucinazioni | 8 / 9 | 89% | ✅ Superato |
| 🏆 TOTALE COMPLESSIVO | Tutti i Dataset Valutati | 46 / 100 | 46% | 🏆 46% Pass |
Protocollo: code_execution, continuation_logprob, cot_generation, letter_logprob · temp 0.0 · seed 42
Impronta di riproducibilità: SHA256-6F3B83EB3171201A
> ⚠️ Misurato su una porzione del dataset, non sulla suite intera: il punteggio non è confrontabile con uno ottenuto sull'intero.
- Data Test:
2026-09-01su motore deterministico SigmaEngine
⏱️ Throughput Hardware e Fasce Consigliate
- Velocità Verificata in Locale:
434.8 tok/ssuNVIDIA GeForce RTX 5070 Ti. - Risposta singola (quello che si sente in chat):
389.1 tok/ssuNVIDIA GeForce RTX 5070 Ti. - Lettura del prompt:
1812 tok/s. - Throughput complessivo durante la valutazione:
434.8 tok/s— piu' richieste in volo insieme, non la velocita' di una risposta singola. - Le velocita' su altro hardware non sono state misurate e non vengono indovinate: dipendono da banda di memoria, quantizzazione, lunghezza del contesto e driver.
⭐ Supporta il Progetto Open Source
Se questo modello ti è utile o vuoi esplorare l'ecosistema completo:
- 🌟 Metti una Stella al repository GitHub: Sigmanih/SigmaStudio
- ❤️ Lascia un Like a questa scheda su Hugging Face
---
Creato e distribuito con il Model Hub di Σ-SIGMA Studio (01/09/2026 19:10)
Run sigmanih/Qwen-Qwen3-0.6B-GGUF-Q6_K with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models