Koshkasa author hub
What's that? The goal : Make a high quality quant of sophosympatheia/Magistry 24B v1.1 using SOTA quant types from ik llama.cpp, allowing the resulting gguf to fit into 16gb VRAM with KV buffer, without KVO, with system overhead. The result : Mixed precision quantization of soph…
Models
7
Downloads
193
Run models locally with guIDE
Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.