Model Intelligence Sheet
Benson-Chen/tether-gguf-models overview
Benson-Chen/tether-gguf-models GGUF downloads for local AI inference with guIDE.
Runs locally from ~235.9 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
46 GGUF files detected
Direct downloads for local inference
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| bonsai-1.7B/Bonsai-1.7B-Q1_0.gguf | GGUF | Q1_0 | 236.8 MB | Download |
| bonsai-1.7B/Bonsai-1.7B-Q2_0.gguf | GGUF | Q2_0 | 441.8 MB | Download |
| bonsai-27B/Bonsai-27B-Q2_0.gguf | GGUF | Q2_0 | 6.67 GB | Download |
| bonsai-4B/Bonsai-4B-Q1_0.gguf | GGUF | Q1_0 | 545.8 MB | Download |
| bonsai-4B/Bonsai-4B-Q2_0.gguf | GGUF | Q2_0 | 1.00 GB | Download |
| bonsai-8B/Bonsai-8B-Q1_0.gguf | GGUF | Q1_0 | 1.08 GB | Download |
| bonsai-8B/Bonsai-8B-Q2_0.gguf | GGUF | Q2_0 | 2.03 GB | Download |
| lfm2-1.2B/LFM2-1.2B-F16.gguf | GGUF | F16 | 2.18 GB | Download |
| lfm2-1.2B/LFM2-1.2B-Q4_K_M.gguf | GGUF | Q4_K_M | 697.0 MB | Download |
| lfm2-1.2B/LFM2-1.2B-TQ2_0.gguf | GGUF | GGUF | 362.5 MB | Download |
| minicpm5-1B/MiniCPM5-1B-F16.gguf | GGUF | F16 | 2.02 GB | Download |
| minicpm5-1B/MiniCPM5-1B-Q4_K_M.gguf | GGUF | Q4_K_M | 656.2 MB | Download |
| minicpm5-1B/MiniCPM5-1B-TQ2_0.gguf | GGUF | GGUF | 436.7 MB | Download |
| qwen3-0.6B/Qwen3-0.6B-Q4_K_M.gguf | GGUF | Q4_K_M | 378.3 MB | Download |
| qwen3-0.6B/Qwen3-0.6B-TQ2_0_Tether.gguf | GGUF | GGUF | 235.9 MB | Download |
| qwen3-1.7B/Qwen3-1.7B-BF16.gguf | GGUF | BF16 | 3.79 GB | Download |
| qwen3-1.7B/Qwen3-1.7B-Q2_K.gguf | GGUF | Q2_K | 741.8 MB | Download |
| qwen3-1.7B/Qwen3-1.7B-Q3_K_M.gguf | GGUF | Q3_K_M | 896.0 MB | Download |
| qwen3-1.7B/Qwen3-1.7B-Q4_0.gguf | GGUF | Q4_0 | 1005.6 MB | Download |
| qwen3-1.7B/Qwen3-1.7B-Q4_K_M.gguf | GGUF | Q4_K_M | 1.03 GB | Download |
| qwen3-1.7B/Qwen3-1.7B-Q4_K_M_imat.gguf | GGUF | Q4_K_M_IMAT | 1.19 GB | Download |
| qwen3-1.7B/Qwen3-1.7B-TQ1_0.gguf | GGUF | GGUF | 533.1 MB | Download |
| qwen3-1.7B/Qwen3-1.7B-TQ2_0.gguf | GGUF | GGUF | 596.1 MB | Download |
| qwen3-1.7B/Qwen3-1.7B-TQ2_0_128.gguf | GGUF | GGUF | 442.0 MB | Download |
| qwen3-4B/Qwen3-4B-BF16.gguf | GGUF | BF16 | 7.50 GB | Download |
| qwen3-4B/Qwen3-4B-Q4_K_M_imat.gguf | GGUF | Q4_K_M_IMAT | 2.33 GB | Download |
| qwen3-4B/Qwen3-4B-TQ2_0_128.gguf | GGUF | GGUF | 1.00 GB | Download |
| qwen3-8B/Qwen3-8B-BF16.gguf | GGUF | BF16 | 15.26 GB | Download |
| qwen3-8B/Qwen3-8B-Q4_K_M_imat.gguf | GGUF | Q4_K_M_IMAT | 4.68 GB | Download |
| qwen3-8B/Qwen3-8B-TQ2_0_128.gguf | GGUF | GGUF | 2.03 GB | Download |
| qwen3.5-0.8B/Qwen3.5-0.8B-BF16.gguf | GGUF | BF16 | 1.41 GB | Download |
| qwen3.5-0.8B/Qwen3.5-0.8B-Q4_K_M.gguf | GGUF | Q4_K_M | 503.1 MB | Download |
| qwen3.5-0.8B/Qwen3.5-0.8B-TQ2_0.gguf | GGUF | GGUF | 333.6 MB | Download |
| qwen3.5-2B/Qwen3.5-2B-BF16.gguf | GGUF | BF16 | 3.52 GB | Download |
| qwen3.5-2B/Qwen3.5-2B-Q4_K_M-imat.gguf | GGUF | Q4_K_M | 1.22 GB | Download |
| qwen3.5-2B/Qwen3.5-2B-Q4_K_M.gguf | GGUF | Q4_K_M | 1.19 GB | Download |
| qwen3.5-2B/Qwen3.5-2B-TQ2_0_128.gguf | GGUF | GGUF | 489.1 MB | Download |
| qwen3.5-2b-allternary/qwen3.5-2b-allternary-TQ2_0_128.gguf | GGUF | GGUF | 489.1 MB | Download |
| qwen3.5-4B/Qwen3.5-4B-BF16.gguf | GGUF | BF16 | 7.85 GB | Download |
| qwen3.5-4B/Qwen3.5-4B-TQ2_0_128.gguf | GGUF | GGUF | 1.05 GB | Download |
| qwen3.5-4b-allternary/qwen3.5-4b-allternary-TQ2_0_128.gguf | GGUF | GGUF | 1.05 GB | Download |
| qwen3.5-9B/Qwen3.5-9B-TQ2_0_128.gguf | GGUF | GGUF | 2.23 GB | Download |
| qwen3.5-vlm-2B/Qwen3.5_VLM-2B-TQ2_0_128.gguf | GGUF | GGUF | 489.1 MB | Download |
| qwen3.5-vlm-4B/Qwen3.5_VLM-4B-TQ2_0_128.gguf | GGUF | GGUF | 1.05 GB | Download |
| qwen3.6-35B-A3B/Qwen3.6-35B-A3B-TQ2_0_128.gguf | GGUF | GGUF | 8.66 GB | Download |
| qwen3.6-35B/Qwen3.6-35B-TQ2_0-router-Q4_K.gguf | GGUF | Q4_K | 8.34 GB | Download |
Model Details
| Model ID | Benson-Chen/tether-gguf-models |
|---|---|
| Author | Benson-Chen |
| Pipeline | — |
| License | — |
| Base model | — |
| Last modified | 2026-08-20T22:55:14.000Z |
Run Benson-Chen/tether-gguf-models with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models