GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

sarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_M overview

Qwen2.5 Coder 7B Instruct GGUF Q3 K M Esta es una versión cuantizada en formato GGUF Q3 K M del modelo original Qwen/Qwen2.5 Coder 7B Instruct . Optimizada esp…

ggufquantizationq3_k_mqwencoderollamabase_model:Qwen/Qwen2.5-Coder-7B-Instructbase_model:quantized:Qwen/Qwen2.5-Coder-7B-Instructlicense:apache-2.0endpoints_compatibleregion:usconversational

Runs locally from ~3.55 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

1 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
qwen-coder-7b-custom-q3_k_m.ggufGGUFQ3_K_M3.55 GBDownload

Model Details

Model IDsarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_M
Authorsarayar
Pipeline
Licenseapache-2.0
Base modelQwen/Qwen2.5-Coder-7B-Instruct
Last modified2026-07-23T04:25:30.000Z

Model README

---

license: apache-2.0

base_model: Qwen/Qwen2.5-Coder-7B-Instruct

tags:

  • gguf
  • quantization
  • q3_k_m
  • qwen
  • coder
  • ollama

---

Qwen2.5-Coder-7B-Instruct GGUF (Q3_K_M)

Esta es una versión cuantizada en formato GGUF (Q3_K_M) del modelo original Qwen/Qwen2.5-Coder-7B-Instruct.

Optimizada especialmente para correr de forma ultra fluida en GPUs con espacio de VRAM ajustado (como las tarjetas de 4 GB de VRAM), manteniendo un equilibrio excelente en tareas de desarrollo de software y programación.

Uso directo en Ollama

Puedes descargar e interactuar con este modelo directamente usando Ollama ejecutando:

ollama run hf.co/sarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_M

Run sarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_M with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models