GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →

tepirale author hub

Model Q2 0.gguf is completely degraded by the quantization. Model Q4 1.gguf shows no quantization problems. MTP active GPU RTX 3090 | 24 GB VRAM Q4 1 use: 18.8 GB VRAM Q5 K S use: 20 GB VRAM MTP active GPU RTX 6000 | 48 GB VRAM max tok | compl tok | time s | tok/s medido | tok/s…

Models
6
Downloads
2,351
tepirale/Ornith-Agents-A1-3.6-35B-A3B-MTP-GGUF
2,346 downloads · image-text-to-text
tepirale/Laguna-S-2.1-DFLASH-GGUF
0 downloads · text-generation

Run models locally with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models