WineryLabs/Winery-Qwen3.5-9B-Atelier-GGUF overview
<p align="center" <a href="https://huggingface.co/WineryLabs" <img src="https://winery api zandy.zocomputer.io/card/m/atelier.svg" alt="Winery Qwen3.5 9B Ateli…
Runs locally from ~8.87 GB disk (12 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| Winery-Qwen3.5-9B-Atelier-Q8_0.gguf | GGUF | Q8_0 | 8.87 GB | Download |
Model Details
| Model ID | WineryLabs/Winery-Qwen3.5-9B-Atelier-GGUF |
|---|---|
| Author | WineryLabs |
| Pipeline | text-generation |
| License | apache-2.0 |
| Base model | Qwen/Qwen3.5-9B,Tesslate/OmniCoder-9B,smirki/tesslate-sft-9b-v3-step-80,OrionLLM/OxCoder-9B,Jackrong/Qwopus3.5-9B-Coder,YangLiu258/Qwen3.5-9B-Exp09-DPO-Html0708,armand0e/Qwen3.5-9B-Coder |
| Last modified | 2026-10-03T11:27:52.000Z |
Model README
---
license: apache-2.0
base_model:
- Qwen/Qwen3.5-9B
- Tesslate/OmniCoder-9B
- smirki/tesslate-sft-9b-v3-step-80
- OrionLLM/OxCoder-9B
- Jackrong/Qwopus3.5-9B-Coder
- YangLiu258/Qwen3.5-9B-Exp09-DPO-Html0708
- armand0e/Qwen3.5-9B-Coder
base_model_relation: merge
library_name: gguf
pipeline_tag: text-generation
tags: [gguf, merge, qwen3.5, winery, llama.cpp, web-design, threejs, frontend, html]
---
<p align="center"><a href="https://huggingface.co/WineryLabs"><img src="https://winery-api-zandy.zocomputer.io/card/m/atelier.svg" alt="Winery Qwen3.5-9B-Atelier" width="100%"/></a></p>
🍷 Winery Qwen3.5 9B · Atelier
A web design + 3D specialist: a region-weighted merge of six Qwen3.5-9B web and coding fine-tunes, built with the Winery fusion compiler. Give it a brief and it writes a complete single-file site: bold landing pages, glassmorphism, scroll animations and live three.js scenes. 9.5 GB at Q8_0.
👉 Atelier's own Space was designed and coded by this model.
Grapes
| Donor | Why it's in |
|---|---|
| Tesslate/OmniCoder-9B | front-end / UI generation |
| smirki/tesslate-sft-9b-v3 | Tesslate UI SFT checkpoint |
| OrionLLM/OxCoder-9B | general coding |
| Jackrong/Qwopus3.5-9B-Coder | Opus-style coding |
| YangLiu258/Qwen3.5-9B-Exp09-DPO-Html0708 | HTML preference-tuned |
| armand0e/Qwen3.5-9B-Coder | coding |
Embeddings and early layers lean on the UI tunes, later layers get more of the coders. All 427 tensors were mixed with zero fallbacks; full recipe in recipe.txt.
Showcase
Unedited one-shot outputs (thinking off, temp 0.6), in showcase/:
| 3D hero (three.js) | Landing page |
|---|---|
Picked over two sibling blends because it was the only one whose three.js scene rendered without errors. It hasn't been benchmarked on MMLU-style tests; it's tuned for front-end work, not trivia.
Use
Turn thinking off for straight code output (chat_template_kwargs: {"enable_thinking": false}) and give it room: full pages run 2-8k tokens.
llama-server -hf WineryLabs/Winery-Qwen3.5-9B-Atelier-GGUF -c 32768
ollama run hf.co/WineryLabs/Winery-Qwen3.5-9B-Atelier-GGUF
Works in llama.cpp (recent builds with Qwen3.5 support), Jan, LM Studio and the Winery app.
Run WineryLabs/Winery-Qwen3.5-9B-Atelier-GGUF with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models