WineryLabs/Winery-Qwen3.8-27B-GrandCru-GGUF overview
<p align="center" <a href="https://huggingface.co/WineryLabs" <img src="https://winery api zandy.zocomputer.io/card/m/27b.svg" alt="Winery Qwen3.8 27B Grand Cr…
Runs locally from ~26.63 GB disk (32 GB+ VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Winery-Qwen3.8-27B-GrandCru-Q8_0.gguf | GGUF | Q8_0 | 26.63 GB | Download |
Model Details
| Model ID | WineryLabs/Winery-Qwen3.8-27B-GrandCru-GGUF |
|---|---|
| Author | WineryLabs |
| Pipeline | text-generation |
| License | other |
| Base model | Qwen/Qwen3.8-27B,ukisai/Swift-1.5-Qwen3.8-27b,Jackrong/Qwopus3.8-27B-Flash,TeichAI/Qwen3.8-27B-Fable-Distill,DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU,huihui-ai/Huihui-Qwen3.8-27B-abliterated,ukisai/Swift-Qwen3.8-27b |
| Last modified | 2026-10-04T23:12:09.000Z |
Model README
---
license: other
license_name: swift-open-license-1.0
license_link: LICENSE
base_model:
- Qwen/Qwen3.8-27B
- ukisai/Swift-1.5-Qwen3.8-27b
- Jackrong/Qwopus3.8-27B-Flash
- TeichAI/Qwen3.8-27B-Fable-Distill
- DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion-735-882-Heretic-Uncensored-NM-DAU
- huihui-ai/Huihui-Qwen3.8-27B-abliterated
- ukisai/Swift-Qwen3.8-27b
base_model_relation: merge
library_name: gguf
pipeline_tag: text-generation
tags: [gguf, merge, qwen3.8, winery, llama.cpp, soup, reasoning, 27b]
---
<p align="center"><a href="https://huggingface.co/WineryLabs"><img src="https://winery-api-zandy.zocomputer.io/card/m/27b.svg" alt="Winery Qwen3.8-27B Grand Cru" width="100%"/></a></p>
🍷 Winery Qwen3.8 27B · Grand Cru
The biggest dense bottle in the cellar: a weighted soup of Qwen3.8-27B and six of its best fine-tunes: Swift 1.5 and Swift 1.0, Qwopus Flash, two Fable-5 distills and Huihui's abliterated build. Every tensor was streamed straight from the donors' bf16 weights and averaged in fp32, then quantised once to Q8_0. 28.6 GB.
Tasting notes
Winery quick eval (0-shot, thinking off; same harness as every Winery model):
| | MMLU | ARC-Challenge | GSM8K | Avg |
|---|---|---|---|---|
| 27B Grand Cru | 76.5 | 98.0 | 81.7 | 85.4 |
| 9B Assemblage | 73.0 | 97.3 | 88.3 | 86.2 |
| 9B Grand Cru | 70.0 | 96.7 | 90.0 | 85.6 |
Grapes (linear soup, fp32 from bf16)
| Donor | Weight |
|---|---|
| Qwen/Qwen3.8-27B (base) | 0.5 |
| ukisai/Swift-1.5-Qwen3.8-27b | 1.0 |
| Jackrong/Qwopus3.8-27B-Flash | 1.0 |
| TeichAI/Qwen3.8-27B-Fable-Distill | 0.9 |
| DavidAU/Qwen3.8-27B-TURBO-Fable-Cold-Fusion | 0.7 |
| huihui-ai/Huihui-Qwen3.8-27B-abliterated | 0.5 |
| ukisai/Swift-Qwen3.8-27b | 0.5 |
All 850 language-model tensors line up across the seven checkpoints. The MTP head is dropped. Config in recipe.json.
Use
Turn thinking off for straight answers (chat_template_kwargs: {"enable_thinking": false}).
llama-server -hf WineryLabs/Winery-Qwen3.8-27B-GrandCru-GGUF -c 32768
ollama run hf.co/WineryLabs/Winery-Qwen3.8-27B-GrandCru-GGUF
Needs about 30 GB of VRAM + RAM combined; llama.cpp's --fit spreads it across a 16 GB GPU and system RAM. Works in llama.cpp (recent builds with Qwen3.8 support), Jan, LM Studio and the Winery app.
Licence
The Swift weights carry the Swift Open License v1.0 (LICENSE); Qwen3.8-27B and the other donors are Apache 2.0 (LICENSE-APACHE-2.0). Attribution in NOTICE.
Copyright 2026 UkisAI. Swift Contribution licensed under the Swift Open License v1.0 (https://huggingface.co/ukisai/Swift-1.5-Qwen3.8-27b/blob/main/LICENSE). Derivative of Qwen3.8-27B, Copyright 2026 Alibaba Cloud, Apache License 2.0.
Run WineryLabs/Winery-Qwen3.8-27B-GrandCru-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models