sarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_M overview
Qwen2.5 Coder 7B Instruct GGUF Q3 K M Esta es una versión cuantizada en formato GGUF Q3 K M del modelo original Qwen/Qwen2.5 Coder 7B Instruct . Optimizada esp…
Runs locally from ~3.55 GB disk (4 GB VRAM class GPUs with llama.cpp / guIDE).
Repository Files & Downloads
| File | Type | Quantization | Size | Link |
|---|---|---|---|---|
| qwen-coder-7b-custom-q3_k_m.gguf | GGUF | Q3_K_M | 3.55 GB | Download |
Model Details
Model README
---
license: apache-2.0
base_model: Qwen/Qwen2.5-Coder-7B-Instruct
tags:
- gguf
- quantization
- q3_k_m
- qwen
- coder
- ollama
---
Qwen2.5-Coder-7B-Instruct GGUF (Q3_K_M)
Esta es una versión cuantizada en formato GGUF (Q3_K_M) del modelo original Qwen/Qwen2.5-Coder-7B-Instruct.
Optimizada especialmente para correr de forma ultra fluida en GPUs con espacio de VRAM ajustado (como las tarjetas de 4 GB de VRAM), manteniendo un equilibrio excelente en tareas de desarrollo de software y programación.
Uso directo en Ollama
Puedes descargar e interactuar con este modelo directamente usando Ollama ejecutando:
ollama run hf.co/sarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_MRun sarayar/Qwen2.5-Coder-7B-Instruct-GGUF-Q3_K_M with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.
Source: Hugging Face · Compare models