GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

Solstice-AI/Qwopus3.8-27B-Flash-GGUF-1M overview

<p align="center" <img src="https://cdn uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice AI Banner"…

ggufsolstice-aiqwenqwen3.8qwopus27bvisionmultimodalmmproj1m-contextlong-contextyarnimage-text-to-textconversationalenzhbase_model:Jackrong/Qwopus3.8-27B-Flashbase_model:quantized:Jackrong/Qwopus3.8-27B-Flashlicense:apache-2.0endpoints_compatibleregion:usimatrix

Runs locally from ~888.0 MB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
image-text-to-text

Repository Files & Downloads

6 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Qwopus3.8-27B-Flash-UD-IQ4_XS-1M.ggufGGUFIQ4_XS13.27 GBDownload
Qwopus3.8-27B-Flash-UD-Q4_K_XL-1M.ggufGGUFQ4_K_XL16.35 GBDownload
Qwopus3.8-27B-Flash-UD-Q6_K_XL-1M.ggufGGUFQ6_K_XL23.56 GBDownload
Qwopus3.8-27B-Flash-UD-Q8_K_XL-1M.ggufGGUFQ8_K_XL29.30 GBDownload
mmproj-BF16.ggufGGUFBF16888.0 MBDownload
speculative/Qwopus3.8-27B-DSpark-Q8_0.ggufGGUFQ8_01.85 GBDownload

Model Details

Model IDSolstice-AI/Qwopus3.8-27B-Flash-GGUF-1M
AuthorSolstice-AI
Pipelineimage-text-to-text
Licenseapache-2.0
Base modelJackrong/Qwopus3.8-27B-Flash
Last modified2026-09-09T01:44:04.000Z

Model README

---

language:

  • en
  • zh

license: apache-2.0

base_model: Jackrong/Qwopus3.8-27B-Flash

tags:

  • solstice-ai
  • qwen
  • qwen3.8
  • qwopus
  • 27b
  • vision
  • multimodal
  • mmproj
  • 1m-context
  • long-context
  • yarn
  • gguf

pipeline_tag: image-text-to-text

---

<p align="center">

<img src="https://cdn-uploads.huggingface.co/production/uploads/67c2e844e0921a5410eec10a/Y5M42dCag2f7Fc6fDtV0Z.jpeg" alt="Solstice-AI Banner" width="100%">

</p>

<h1 align="center">Qwopus3.8-27B-Flash-1M (Unsloth Dynamic v3.0 GGUF)</h1>

<h3 align="center">Official Solstice-AI Quantization &bull; Native-Like 1M Context Window &bull; Full Multimodal Vision &bull; Zero Command Flags Required</h3>

<p align="center">

<img src="https://img.shields.io/badge/org-Solstice--AI-blueviolet" alt="Solstice-AI">

<img src="https://img.shields.io/badge/license-Apache%202.0-blue" alt="License">

<img src="https://img.shields.io/badge/format-GGUF-orange" alt="Format">

<img src="https://img.shields.io/badge/precision-Unsloth%20Dynamic%20v3.0-yellow" alt="Precision">

<img src="https://img.shields.io/badge/context-1M%20Tokens-purple" alt="Context">

<img src="https://img.shields.io/badge/arc--c-735-brightgreen" alt="ARC-C">

</p>

---

Model Overview

Solstice-AI/Qwopus3.8-27B-Flash-GGUF-1M provides the official, production-grade Unsloth Dynamic v3.0 GGUF release of Qwopus3.8-27B-Flash with a native-behaving 1,048,576-token (1M) context window.

Unsloth Dynamic v3.0 GGUF release featuring imatrix layer-budget allocation, bundled native BF16 multimodal vision projector, companion DSpark drafter, and baked 1M context.

Key Specifications

| Attribute | Specification |

| :--- | :--- |

| Base Model | Jackrong/Qwopus3.8-27B-Flash |

| Architecture | Qwen3.5 / Qwopus Conditional Generation with Multimodal Vision |

| Quantization Methodology | Unsloth Dynamic v3.0 (Selective high-precision budget layers) |

| Context Window | 1,048,576 tokens (1M native context) |

| Multimodal Vision | Standalone native BF16 projector (mmproj-BF16.gguf) |

| Bundled Drafter | Companion 27B DSpark speculative drafter in speculative/ |

| Target Engines | llama.cpp, Ollama, LM Studio, Unsloth |

---

Benchmark Highlights & Validation

Evaluated under the standardized benchmark harness:

| Benchmark Suite | Discipline | Qwopus3.8-27B-Flash (1M) | Claude Opus 4.6 Max | GPT-4o |

| :--- | :--- | :---: | :---: | :---: |

| SWE-bench Pro | Agentic Software Engineering | 61.7% | 53.4% | 48.9% |

| LiveCodeBench v6 | Algorithmic Problem Solving | 90.3% | 88.8% | 72.8% |

| QwenSWEBench | Complex Architecture Refactoring | 79.0% | 63.8% | 61.2% |

| OSWorld-Verified | Desktop & Operating System Automation | 84.3% | 72.7% | 58.7% |

| ARC-C (Challenge) | Frontier Scientific Reasoning | 735 (8-Bit) / 719 (4-Bit) | ~710–720 | 63.8% |

| Long-Context Needle | 256K &rarr; 1M Tokens Retrieval | 100% (Bit-Exact) | Pass | Pass |

---

Attribution & Acknowledgments

Run Solstice-AI/Qwopus3.8-27B-Flash-GGUF-1M with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models