GraySoft
Projects Models Compare Cloud benchmarks FAQ Download guIDE →
Model Intelligence Sheet

ilintar/Agents-A1-GGUF overview

Optimized with my branch's custom auto tensor type , custom made recipes for 3.77, 4.02 and 4.27 bpw element gamma=0.25, tuned for MoE — recovers the bits that…

ggufbase_model:InternScience/Agents-A1base_model:quantized:InternScience/Agents-A1license:apache-2.0endpoints_compatibleregion:usimatrixconversational

Runs locally from ~15.21 GB disk (16 GB VRAM class GPUs with llama.cpp / guIDE).

Downloads
0
Likes
0
Pipeline
Author

Repository Files & Downloads

3 GGUF files detected
Direct downloads for local inference
FileTypeQuantizationSizeLink
Agents-A1-IQ3_M.ggufGGUFIQ3_M15.21 GBDownload
Agents-A1-IQ4_XS.ggufGGUFIQ4_XS17.23 GBDownload
Agents-A1-Q3_K_M.ggufGGUFQ3_K_M16.22 GBDownload

Model Details

Model IDilintar/Agents-A1-GGUF
Authorilintar
Pipeline
Licenseapache-2.0
Base modelInternScience/Agents-A1
Last modified2026-07-09T22:57:18.000Z

Model README

---

license: apache-2.0

base_model:

  • InternScience/Agents-A1

---

Optimized with my branch's custom auto-tensor-type, custom-made recipes for 3.77, 4.02 and 4.27 bpw

(element-gamma=0.25, tuned for MoE — recovers the bits that plain size-weighting over-spends on the

rarely-activated experts).

Since HF doesn't recognize custom bpw tags, I've tagged them with:

  • IQ3_M: 3.77bpw
  • Q3_K_M: 4.02bpw
  • IQ4_XS: 4.27bpw

Note that the quant types are only aliases for the size and do not correspond to the actual quant types used.

Converted with --no-mtp, so the multi-token-prediction head is excluded — these are standard-inference GGUFs.

Run ilintar/Agents-A1-GGUF with guIDE

Download guIDE — the AI-native code editor with local LLM inference and 69 built-in tools.

Download guIDE → · Browse 524k+ models · Compare models

Source: Hugging Face · Compare models